Slack engineering
Read post

Advancing Our Chef Infrastructure: Safety Without Disruption

Slack improved their Chef infrastructure safety by splitting a single production environment into six isolated buckets (prod-1 through prod-6) mapped to availability zones, implementing a release train model with staggered rollouts. They built Chef Summoner, a service that triggers Chef runs based on S3 signals rather than fixed cron schedules, reducing blast radius during deployments. The approach avoided disruptive migration to Policyfiles while achieving safer deployments. Changes now take longer to propagate but provide time to catch issues before full rollout. A fallback cron job ensures Chef runs every 12 hours even if Summoner fails, maintaining compliance.

    #aws#devops#infrastructure#cicd#chef
Oct 23, 2025•15m read time•From slack.engineering
Post cover image
Table of contents
Splitting Chef EnvironmentsWhat changed in the way we trigger Chef?What’s Next?
1.4K Impressions
Slack engineering's image
Slack engineering

The Slack Blog serves as a resource for teams and developers looking to make the most out of Slack, ...

98 Followers

•

239 Upvotes

Would you recommend this post?

Copy link
WhatsApp
Facebook
X
New Squad
  • © 2026 Daily Dev Ltd.
  • Guidelines
  • Explore
  • Tags
  • Sources
  • Squads
  • Leaderboard