Waste Ratio
The proportion of tokens consumed that produce no useful output. The metric behind Token Bleed and Token Burn.
Waste Ratio is the aggregate measurement of all token waste in a system. It includes: tokens spent on re-onboarding (Token Bleed), tokens spent re-processing old conversation history (Token Burn), tokens spent on Reasoning Load from poor context, and tokens spent on failed attempts that had to be retried. Waste Ratio gives you a single number that captures how efficiently your system uses its token budget. Most teams are shocked to discover their Waste Ratio is 40-60% — meaning half their AI spending produces no value.
More in Measurement
Context Efficiency
Useful retained state per token consumed.
Token ROI
The measurable value produced per token consumed. The business metric for Context Efficiency.
Context Utilization Rate
The percentage of input context that the model actually uses for its output. Low rate = waste.
Recall Cost Ratio
The cost of retrieving stored context vs. the cost of re-generating it from scratch.
Iteration Velocity
The speed at which a human-AI pair can move from idea to tested output.
Recovery Cost
The time and tokens required to get a session back on track after a failure — State Loss, Execution Hallucination, or Plan Drift.
Quality-Per-Token
A measure of output quality relative to tokens consumed. Distinct from Token ROI (value) — this measures correctness, completeness, and relevance per unit of cost.