Counts are provider-reported across answer, diagnostic and judge calls. Each case sends different evidence, so reuse comes mainly from shared instructions, and providers cache only prefixes above a model-specific minimum length. {% if not cache.tracked %}This run predates per-call cache tracking; the hit rate and writes were not recorded.{% endif %}
Usage is reported by the provider. Missing prices or usage mean the cost is unknown. Compare models & rerankers for per-call measurements.
| Purpose | Input tokens | Cached input | Cache writes | Cache hit rate | Output tokens | API cost |
|---|---|---|---|---|---|---|
| {{ label }} | {{ usage.get(prefix ~ '_prompt_tokens', '—') }} | {{ usage.get(prefix ~ '_cached_tokens', '—') }} | {{ usage.get(prefix ~ '_cache_write_tokens', '—') }} | {% if reported %}{{ '%.1f%%'|format(100 * usage.get(prefix ~ '_cache_hit_calls', 0) / reported) }} ({{ usage.get(prefix ~ '_cache_hit_calls', 0) }}/{{ reported }}){% else %}—{% endif %} | {{ usage.get(prefix ~ '_completion_tokens', '—') }} | {% set cost = usage.get('lane_cost_usd', {}).get(lane) %}{% if cost is not none %}${{ '%.6f'|format(cost) }}{% else %}Unknown{% endif %} |
{{ usage|tojson(indent=2) }}