The scoreboard, honestly: the hard targets, how often each one is actually looked at,
and the quiet human signals that never make it onto a dashboard.
Mean Time To Resolution (MTTR) for Major Incidents (P1)
The average time it takes from a P1 incident being declared to the service being fully restored.
Target · < 45 minutesIf we had three P1 incidents this month, resolved in 30, 50, and 40 minutes respectively, your average MTTR would be 40 minutes, hitting the target.
Reduction in Recurring Incidents
The percentage decrease in incidents that are direct repeats of previously resolved issues within a given period, showing effective problem management.
Target · 15% reduction quarter-over-quarterIf Q1 saw 20 recurring incidents, and Q2 saw 17 (a 15% drop), you're on track. This means your solutions are actually sticking.
Knowledge Base Articles Authored/Updated
The number of new or significantly updated knowledge base articles, runbooks, or SOPs you've created to help the wider team.
Target · > 5 per quarterYou've written a new troubleshooting guide for a common VoIP issue and updated three outdated switch configuration procedures. That's four, so you'd need one more to hit the target.
Project Completion Rate for Operational Improvements
The percentage of small to medium operational improvement projects (e.g., monitoring enhancements, automation scripts) that you lead and complete on time.
Target · > 90%You were tasked with implementing a new alert correlation rule in SolarWinds and automating a daily health check script. If both are done on schedule, that's 100%.
Technical Solution Adoption and Effectiveness
How well the solutions you design and implement are received and used by the team, and if they genuinely solve the underlying problem.
- Team members consistently use your new runbooks
- the number of incidents related to the problem you addressed drops significantly
- positive feedback from your manager and peers on your problem-solving approach.
Team Mentorship and Development
How effectively you guide, teach, and uplift the junior technicians reporting to you, helping them grow their skills and confidence.
- Junior team members are able to resolve more complex issues independently
- positive feedback from your direct reports in 1-to-1s
- your manager observes you providing constructive feedback and guidance during incident reviews.
Proactive Problem Identification and Prevention
Your ability to spot potential network issues before they become major outages, often by analysing trends or identifying weaknesses.
- You present a proposal to fix a potential single point of failure before it causes an outage
- you identify a pattern of minor errors that, left unchecked, would become a major problem
- your manager notes your foresight in preventing issues.
Crisis Communication and Leadership
How clearly and calmly you communicate during major incidents, providing leadership and technical direction.
- Stakeholders praise your clear updates during an outage
- you effectively delegate tasks and coordinate efforts on an incident bridge
- your manager trusts you to lead critical incident calls.