Triage why a service or system cannot meet target QPS, throughput, concurrency, or job-processing rate by identifying the actual saturated layer (CPU, memory, IO, network, DB, or concurrency limit). Use when scaling workers/instances hasn't fixed a throughput ceiling. Produces a throughput bottleneck triage with a bottleneck hypothesis, evidence table, and safe experiments.