Input tokens served from the provider's prompt cache.
Input tokens written into the provider's prompt cache.
Best-effort USD estimated from list prices at call time.
Prompt (input) tokens.
Number of LLM calls billed to this model.
The resolved model these totals are attributed to.
Completion (output) tokens.
Reasoning tokens, where the provider reports them separately.
One model's slice of a KB's graph LLM usage, keyed by the resolved model name. Token counts are the provider-reported ground truth;
estCostUsdis estimated from model list prices at call time.