OptionalbyPer-model split of the totals above (LLM calls only), present when any usage was recorded. The per-model figures sum to the aggregate; embedding usage is excluded from both.
Input tokens served from the provider's prompt cache.
Input tokens written into the provider's prompt cache.
Best-effort USD estimated from list prices at call time.
Prompt (input) tokens.
Number of LLM calls.
Completion (output) tokens.
Reasoning tokens, where the provider reports them separately.
Aggregated LLM token usage (and best-effort cost) of graph work on a KB. Token counts are the provider-reported ground truth;
estCostUsdis estimated from model list prices at call time — recompute from the tokens for exact accounting.