Markdown · Illustrative example
Benchmark Gap Investigation — Squad-Name — Apr–Jun 2026
Scope and coverage
90 days · all repositories · benchmark source: DevStats peer cohort · 11 of 12 metrics covered.
Benchmark signals to inspect
| Metric | Tier | Evidence | Why inspect |
|---|---|---|---|
| PR Cycle Time | Medium | 6.8 days; peer median 4.2 | Review stage represents 46% of elapsed time |
| Change Failure Rate | Low | 18%; peer median 11% | 5 of 28 deployments were followed by a failure |
Diagnostic read
- Review time increased from 2.1 to 3.1 days while PR size remained stable; this is compatible with queueing and deserves context.
- Deployment volume stayed near 9 per month, while failed changes rose from 2 to 5; this does not identify a technical cause.
Interpretation limits
- Benchmark tiers are external reference points, not a team score or target.
- One repository had incomplete incident linkage for 12 days.
To understand the context
- Where did review queues concentrate during May and June?
- What changed in incident classification or release conditions during the period?
Evidence used
- Benchmarks — identify covered tiers
- PR Cycle Time — validate review-stage concentration
- DORA Metrics — validate stability signal