How to spend a latency budget across a multi-hop AI request, and how to defend p99 when one hop misbehaves.
Engineers with a latency SLO on an AI feature.