The scoreboard, honestly: the hard targets, how often each one is actually looked at,
and the quiet human signals that never make it onto a dashboard.
Mean Time To Resolution (MTTR) for P1/P2 Incidents
The average time it takes from an incident being reported to it being fully resolved and services restored for our most critical issues.
Target · Reduce MTTR for P1 incidents from 4 hours to 2.5 hours within 12 months; for P2 incidents, from 8 hours to 5 hours.If our current P1 MTTR is 4 hours, and you help us bring it down to 2.8 hours, that's a huge win. This means less business disruption, faster recovery, and happier users. We'll track this directly from our ITSM platform, looking at the full lifecycle of each major incident.
Successful Change Implementation Rate
The percentage of planned IT changes (e.g., software deployments, infrastructure upgrades) that are implemented without causing a new incident or service degradation.
Target · Maintain a >99% successful change implementation rate for all production changes.If we have 100 planned changes in a quarter and only one causes an outage or significant issue, that's a 99% success rate. You'll be looking at the changes you've approved and the processes your team manages, making sure the 'change freeze' periods are respected and that proper testing happens. A successful change means no unexpected P1s.
Problem Resolution Rate for Recurring Incidents
The percentage of identified recurring problems (i.e., multiple incidents with the same root cause) that have a permanent fix implemented and verified.
Target · Reduce the number of recurring P1/P2 incidents by 20% year-on-year through effective problem management.Say we had a recurring database connectivity issue that caused three P2 incidents in Q1. If you lead the problem investigation, get engineering to implement a permanent fix, and we see no more incidents from that root cause in Q2, that counts as a successful problem resolution. It's about stopping the same problems from hitting us again and again.
Service Level Agreement (SLA) Adherence for Critical Services
Ensuring that our key IT services consistently meet their defined availability, performance, and support response targets.
Target · Achieve >99.9% availability for Tier 1 business applications and >95% adherence to support response SLAs for P3/P4 tickets.If our core CRM system is meant to be up 99.9% of the time, and it drops to 99.8% due to an incident, you'll be on the hook to explain why and put a plan in place to fix it. This isn't just about incident response; it's about the overall health of our services. You'll be looking at the numbers, identifying trends, and pushing for improvements.
Stakeholder Confidence During Incidents
How effectively you manage communications and expectations with senior leadership and affected business units during major incidents, ensuring they feel informed and confident in the recovery process.
- Positive feedback from business unit leads and executives during post-incident reviews
- being proactively sought out for updates during crises
- clear, concise, and timely executive communications (no jargon, just facts and actions)
- ability to de-escalate tension and maintain focus on resolution.
Quality of Post-Incident Reviews (PIRs) / Blameless Post-mortems
The thoroughness and effectiveness of your incident analysis, identifying true root causes, and ensuring actionable preventative measures are documented and assigned.
- Comprehensive PIR reports that clearly explain the incident timeline, impact, root cause, and preventative actions
- evidence of follow-up on assigned actions
- feedback from engineering teams that PIRs are constructive and lead to real improvements, not just blame
- a culture of learning from mistakes without fear of reprisal.
Process Improvement Adoption and Impact
Your ability to design, champion, and embed new or improved service delivery processes (e.g., change management, problem management) that genuinely reduce friction and improve outcomes.
- Observable improvements in process efficiency (e.g., faster change approvals, fewer process bottlenecks)
- positive feedback from teams using the new processes
- clear documentation and training materials for new processes
- measurable reduction in incidents or service degradation attributable to process changes.
Team Mentorship and Development
Your effectiveness in guiding and developing the Service Delivery Analysts and Specialists who report to you, helping them grow their skills and take on more complex challenges.
- Evidence of regular 1:1s and performance feedback
- successful delegation of tasks that stretch team members
- positive feedback from direct reports on their development
- demonstrable improvement in team members' technical and soft skills over time
- successful onboarding of new team members who quickly become productive.