Context Utilization Rate
The percentage of input context that the model actually uses for its output. Low rate = waste.
Context Utilization Rate answers a question most teams never ask: of all the tokens we send to the model, how many actually influence the output? If you send 20,000 tokens of context and the model's output is based primarily on 3,000 of them, your Context Utilization Rate is 15%. The other 85% is dead weight — tokens you paid for that the model processed but didn't meaningfully use. Low utilization is a symptom of Prompt Bloat and poor Selective Recall.
More in Measurement
Context Efficiency
Useful retained state per token consumed.
Token ROI
The measurable value produced per token consumed. The business metric for Context Efficiency.
Recall Cost Ratio
The cost of retrieving stored context vs. the cost of re-generating it from scratch.
Waste Ratio
The proportion of tokens consumed that produce no useful output. The metric behind Token Bleed and Token Burn.
Iteration Velocity
The speed at which a human-AI pair can move from idea to tested output.
Recovery Cost
The time and tokens required to get a session back on track after a failure — State Loss, Execution Hallucination, or Plan Drift.
Quality-Per-Token
A measure of output quality relative to tokens consumed. Distinct from Token ROI (value) — this measures correctness, completeness, and relevance per unit of cost.