Reliability Strategy Session

Mat Szymczyk

profile
Reliability Strategy Session
profile
200
60 mins

Most organizations wait until production fires are burning before thinking about reliability. By then, fixing problems is 10x more expensive and disruptive.

In this 60-minute strategy session, we'll explore how to "shift left" - bringing reliability thinking into your requirements, architecture, and design phases. Drawing from my 12+ years of DevOps/SRE experience (including optimizing global Kubernetes infrastructure at Shell), I'll help you identify:

  • Where reliability decisions should happen in your current development process
  • Critical reliability risks in your architecture or approach that are cheaper to address now
  • Quick wins you can implement immediately to improve system resilience
  • The right SRE practices for your team's maturity level and constraints
  • How to avoid common reliability mistakes I've seen across enterprise and startup environments

This isn't a generic best practices lecture - it's a collaborative conversation tailored to your specific technology stack, team structure, and business constraints.


Best for:

  • Startups building their first production systems
  • Scale-ups experiencing growing pains
  • Engineering leaders wanting to mature their reliability practices
  • Teams tired of firefighting who want to prevent fires instead


What to prepare before our call:

  • Brief description of your system/architecture (sent 24hrs before)
  • Specific reliability concerns or questions you're facing
  • Willingness to challenge assumptions and think differently


What you'll get:

  1. Clear action items you can implement this week
  2. Strategic direction for building reliability into your process
  3. Honest assessment of what matters now vs. what can wait