Model p99 latency rises while average latency is stable. How would you find the bottleneck?
AceStack AI
Tech Screening
Inference latency
What this task practices
Inference latency is a tech screening interview exercise that trains prompt interpretation, explicit assumptions, a concrete response, and a clear explanation of tradeoffs. The catalog marks it as medium difficulty. It focuses on Debugging, Tradeoffs, Technical Reasoning. The signed-in workspace provides the tools for the round and evaluates the attempt against task-specific criteria. Reference solutions, hidden checks, evaluator instructions, and candidate work remain private.
Sign in to startThe workspace and evaluation open after sign in.