Accelerate was published four years ago and understanding of the 'four DORA metrics' has matured. Štěpán Davidovič wrote a good piece Incident Metrics in SRE [1] where he pulls apart the Mean Time to Recovery Metric. I've written an overview of the DORA metrics and what to use instead - SLOs & SLIs [2]
[1] https://web.archive.org/web/20210917022148/https://static.go...
[2] https://isthisit.nz/posts/2022/state-of-the-dora-devops-metr...
I would start with “can you even measure this”. Just getting something like this instrumented for your team is probably going to drive improvements.
Then you can put it on a Shewhart/PDCA Cycle driven primarily by the team itself through its retrospective and prioritization process. “Has this metric slipped? Do we feel like there are blocks in the deploy pipeline? Are those blocks more important than other work right now? What changes does the team feel most invested in?” Fostering that conversation is really important.
Grabbing these metrics, poorly measuring them, and then trying to hold a team accountable to some arbitrary, external, non-contextual measure is a sure way to kill morale and productivity, hinder future leadership efforts, and waste time and resources and non-priorities.
Of course, doing nothing will accomplish the same.
We do this kind of work with hospitals and their process metrics. Our measurements are always based on raw data, never self-reported metrics. Frequently, team’s self-reported perceptions prior to measurement is the most telling thing. When I worked in k-12 Ed, the same was true with teachers and their students: they perceived their students as great writers, but they were often much worse when measured quantitatively by normalized graders. What is important in both cases isn’t blindly pushing to improve those metrics, but developing an understanding of why that perceived difference exists, why the numbers are the way they are, and ultimately, whether and what kinds of differences you can make. In the case of hospitals, there are some units/teams where you don’t want high performance numbers, because it might lead to higher error rates. In Ed, there might be limits to what you can do as a single-subject teacher focused on a single year in a group of children’s lives. But you can understand each groups challenges and strengths a little better and maybe adapt accordingly.
Blindly pushing top-line metrics is not the same as blindly doing the same thing with no measurement or learning. Both are bad.
https://waldencroft.com/gaming-the-system-the-unintended-con...
If you want to be successful, you have to focus on intrinsic motivations of these kinds of staff. By and large, they do not respond effectively to extrinsic motivations. They may respond "positively", but not necessarily "effectively". It's difficult, though, in a world where any non-quantifiable evaluation is rightly viewed with incredible suspicion. It forces managers and leaders to walk a hard line.