The pathway
How you actually get there, here
How you become one varies far more by country than what one does. This is the UK route. Most people take one of these ways in; the right one depends on where you're starting from.
- 1
Senior Service Delivery Analyst / Problem Manager
3-5 years in a senior individual contributor roleSkills to master
- Leading minor incidents, performing deep-dive root cause analysis, owning specific service improvement plans, mentoring junior team members, and building strong relationships with development teams.
You're ready to move on when
- Consistently leads resolution of complex P3/P4 incidents independently.
- Successfully identifies and resolves recurring problems, showing measurable impact.
- Receives positive feedback on mentorship and cross-functional collaboration.
- Proactively identifies areas for process improvement and proposes solutions.
- 2
Technical Lead / Team Lead (within an Operations team)
2-4 years leading a small technical operations teamSkills to master
- Managing a small team's workload, guiding technical troubleshooting, implementing automation scripts, ensuring operational readiness for new deployments, and contributing to on-call rotations.
You're ready to move on when
- Demonstrates strong technical leadership and problem-solving skills within their domain.
- Effectively manages team performance and development.
- Successfully delivers operational projects on time and within scope.
- Is the go-to person for complex technical issues within their area.
- 3
Site Reliability Engineer (SRE)
4-6 years as a hands-on SRESkills to master
- Defining and implementing SLOs/SLIs, building robust monitoring and alerting systems, automating infrastructure and operational tasks (IaC), participating in blameless post-mortems, and a deep understanding of system resilience.
You're ready to move on when
- Has significantly contributed to improving service reliability and reducing toil.
- Demonstrates expertise in observability, automation, and incident prevention.
- Is comfortable leading technical discussions and driving consensus on reliability improvements.
- Has a strong understanding of the business impact of system stability.