ImprovementBuild ·
Latency, broken into four numbers

What it does
Message Insights splits response time into Pickaxe latency, model time-to-first-token, total time-to-first-token, and generation time.
Why it matters
One aggregate number tells you an agent is slow. These four tell you whether to trim the knowledge budget, switch models, or shorten the output.
