Skip to main content
The Bike Shed

445: Working Iteratively

40 min episode · 2 min read
·
Joelle Kenville,Stephanie Min

Episode

40 min

Read time

2 min

Topics

Productivity, Leadership, Software Development

AI-Generated Summary

Key Takeaways

  • Blameless Post-Mortems: Effective incident retrospectives focus on system-level failures rather than individual mistakes, examining monitoring gaps, deployment timing decisions, and response processes to build resilient engineering cultures where teams learn from production incidents without fear.
  • Deployment Friction Costs: Shared staging environments and slow CI pipelines create per-iteration costs that incentivize bundling changes together. Teams working across time zones with mandatory QA approval face twenty-four hour turnaround times, making developers combine multiple features into single deployments.
  • PR Size and Review Speed: Code review time scales non-linearly with PR size—doubling lines of code can triple review effort. Small, frequent PRs create virtuous cycles where reviewers spend ten minutes between tasks, while large PRs require scheduled hour-long sessions that delay feedback.
  • Iterative Communication Benefits: Shipping one PR per day enables developers to provide specific progress updates about refactoring steps and technical discoveries, replacing vague status reports. Teams gain visibility into work-in-progress, reducing context-loading overhead when switching between tasks and maintaining development momentum.

What It Covers

Stephanie Minn and Joelle Kenville examine the technical and social factors that enable or prevent teams from adopting iterative development practices, including deployment friction, code review culture, and organizational incentives that shape developer behavior.

Key Questions Answered

  • Blameless Post-Mortems: Effective incident retrospectives focus on system-level failures rather than individual mistakes, examining monitoring gaps, deployment timing decisions, and response processes to build resilient engineering cultures where teams learn from production incidents without fear.
  • Deployment Friction Costs: Shared staging environments and slow CI pipelines create per-iteration costs that incentivize bundling changes together. Teams working across time zones with mandatory QA approval face twenty-four hour turnaround times, making developers combine multiple features into single deployments.
  • PR Size and Review Speed: Code review time scales non-linearly with PR size—doubling lines of code can triple review effort. Small, frequent PRs create virtuous cycles where reviewers spend ten minutes between tasks, while large PRs require scheduled hour-long sessions that delay feedback.
  • Iterative Communication Benefits: Shipping one PR per day enables developers to provide specific progress updates about refactoring steps and technical discoveries, replacing vague status reports. Teams gain visibility into work-in-progress, reducing context-loading overhead when switching between tasks and maintaining development momentum.

Notable Moment

Stephanie describes running a large database migration that caused a weekend production incident affecting many customers. The retrospective revealed how multiple system failures compounded, but the team maintained a blame-free culture focused on preventing future occurrences through improved monitoring and deployment practices.

Know someone who'd find this useful?

Episode Transcript

This episode is brought to you by WorkOS. If you're building a b two b ass app, at some point, your customers will start asking for enterprise features like single sign on, skim, provisioning, role based access control, and audit trails. That's where WorkOS comes in. With ease to use and flexible APIs that help you ship enterprise features on day one without slowing down your core product development. Today, some of the hottest startups in the world are already powered by WorkOS, including ones you probably know, like Perplexity, Vercel, Jasper, and Webflow. WorkOS also provides a generous free tier of up to 1,000,000 monthly active users for its user management solution, making it the perfect authentication and authorization solution for growing companies. It comes standard with rich features like social logins, bot protection, MFA, roles, and permissions, and more. If you're currently looking to build SSO for your first enterprise customer, you should consider using Work OS. Integrate in minutes and start shipping enterprise plans today. Check it all out at workos.com. That's workos.com. Hello, and welcome to another episode of the Bike Shed, a weekly podcast from your friends at Thoughtbot about developing great software. I'm Joelle Kenville. And I'm Stephanie Min. And together, we're here to share a bit of what we've learned along the way. So Stephanie, what's new in your world? So I caused an incident at work last week. So that was not totally new because I have before, but I did wanna bring it up because it happens, and it's just part of our jobs. Did you delete the prod database? No. Nothing like that, but it was still bad. It was it had, like, a very large customer impact. I was running a really big planned migration, and I think we had made some assumptions about the relative risk and safety of it. Specifically, you know, I had set this up to not kind of thinking that it would not involve, like, impacting existing records, but just kind of buried deep in the code some kind of weird emergent behavior that I wasn't expecting did happen. And that had another strange side effect that wasn't kind of handled correctly somewhere else in the system, and it happens. So, yeah, I I think that what I wanted to mention specifically about the incident though was that I had we had a retrospective on it a few days afterwards where the people who were involved with like triaging and resolving the incident because it wasn't it wasn't me, it happened over the weekend and the folks who were on call had to deal with it. All of us got together and just talked about the facts of what happened and then areas for improvement specifically at a system level. And one of the things that I really, really appreciated about this retrospective was how blame free it was and, yeah, just not kind of calling any individual out. Right? It was a system …

Get the full transcript (7,307 words) + summary by email — free

One-time email with the complete transcript and AI summary of this episode. No account needed.

One email, no spam. We’ll also show you what SignalCast does.

Browse all The Bike Shed transcripts →

You just read a 3-minute summary of a 37-minute episode.

Get The Bike Shed summarized like this every Monday — plus up to 2 more podcasts, free.

Pick Your Podcasts — Free

Keep Reading

More from The Bike Shed

We summarize every new episode. Want them in your inbox?

Similar Episodes

Related episodes from other podcasts

Explore Related Topics

This podcast is featured in Best Cybersecurity Podcasts (2026) — ranked and reviewed with AI summaries.

Read this week's Software Engineering Podcast Insights — cross-podcast analysis updated weekly.

You're clearly into The Bike Shed.

Every Monday, we deliver AI summaries of the latest episodes from The Bike Shed and 192+ other podcasts. Free for one show.

Start My Monday Digest

No credit card · Unsubscribe anytime