Performance — Professional¶
Linux perf samples hardware and kernel events; eBPF observes scheduling and I/O; JVM Flight Recorder captures runtime events with low overhead; continuous profilers aggregate stack samples across fleets. At 10× load, queueing and pools dominate; at 100×, memory locality, data movement, and fleet economics become architectural.
Design and operations checklist¶
- Define user-visible budgets and representative workloads.
- Correlate latency with saturation and profiles.
- Preserve benchmark reproducibility and variance.
- Test failure and recovery load.
- Gate meaningful regressions and review false positives.
- Track cost per useful operation.
Test yourself¶
- Design performance governance for a polyglot fleet.
- How does coordinated omission corrupt load results?
- Which hardware counters explain cache-bound code?
- When is a performance regression acceptable?
Further reading¶
- Brendan Gregg, Systems Performance.
- Gil Tene on latency and coordinated omission.
- Neil Gunther, Guerrilla Capacity Planning.