o10Reviewed 2026-09-30

The Quality-Cost Frontier Index

What it actually costs, this month, to clear a quality bar. Measured across real routed traffic and tune runs, per task type.

Research · Methodology preview · September 2026

SummaryKey takeaways

What you need to know

Start with the core questions, then examine the examples and tradeoffs below.

What does it cost to run AI at a given quality bar?

The Quality-Cost Frontier Index answers that per task type: the cheapest model that clears a stated quality bar (0.74 / 0.86 / 0.93) and its measured cost per 1K calls, with observation and org counts on every real row.

Methodology preview. ILLUSTRATIVE rows until the first live period publishes.

What is never collected for this index?

Prompt content. The index is metrics-only with k-anonymity minimums on observations and orgs before a cell publishes.

No prompts · k-anonymity thresholds · public reference

How does this relate to o10’s product?

o10 holds your quality floor in the request path and routes each call to the cheapest model that clears it. The index is the public, aggregatable expression of that underwriting. Savings are the consequence of clearing the bar cheaper, not a claim without proof.

01Deep dive

Methodology

How a cell enters the index, and what never does.

Quality is holdout-scored on customer-supplied examples (Tune) and enforced via eval-gated live routing. A model is listed as cheapest-clearing only when it empirically clears the stated quality bar.

Publishing requires k-anonymity minimums: both observation_count and org_count must clear internal thresholds before a cell appears. Prompt content is never collected for this index.

Costs are reported as USD per 1K calls for the clearing model at that task type × quality bar. Periods are calendar months unless the API specifies otherwise.

  • Holdout-scored evals on customer-supplied examples
  • Eval-gated live routing receipts
  • K-anonymity: observation + org thresholds
  • No prompt content collected for the index
SourceMethodology

Source: https://app.o10.io/api/frontier-index. Page resolves live fetch → last cached period → ILLUSTRATIVE methodology preview. Never present sample rows as measured.

IndexMethodology preview · September 2026

Task type × quality bar

Cheapest clearing model and cost per 1K calls . methodology preview. Every sample row is tagged ILLUSTRATIVE and is not measured traffic.

ILLUSTRATIVE methodology preview, not measured · quality bars 0.74 / 0.86 / 0.93
Task type Quality bar Cheapest clearing model $/1K calls n (obs) Orgs Label
classification 0.74 glm-class (illustrative) $0.420 - - ILLUSTRATIVE
classification 0.86 llama-class (illustrative) $1.15 - - ILLUSTRATIVE
classification 0.93 sonnet-class (illustrative) $4.80 - - ILLUSTRATIVE
extraction 0.74 mini-class (illustrative) $0.680 - - ILLUSTRATIVE
extraction 0.86 haiku-class (illustrative) $1.90 - - ILLUSTRATIVE
extraction 0.93 gpt-class (illustrative) $6.20 - - ILLUSTRATIVE
rag_qa 0.74 open-weight-class (illustrative) $0.950 - - ILLUSTRATIVE
rag_qa 0.86 mid-tier-class (illustrative) $2.40 - - ILLUSTRATIVE
rag_qa 0.93 frontier-class (illustrative) $9.10 - - ILLUSTRATIVE

How to cite: o10 Quality-Cost Frontier Index, September 2026. Source page: www.o10.io/research/quality-cost-frontier-index. Distribution: https://app.o10.io/api/frontier-index.

FAQFrequently asked questions

Common questions

What is the Quality-Cost Frontier Index?

A privacy-safe monthly dataset from o10: for each task type and quality bar, the cheapest model that empirically clears the bar and its cost per 1K calls, with observation_count and org_count. It is designed to be cited, not a marketing estimate.

What data does o10 collect for the index?

Metrics only: task type, quality bar, clearing model, cost per 1K calls, and k-anonymized observation/org counts. No prompt content is collected for this index. Rows publish only when observation and org thresholds clear.

How does o10 measure quality for the index?

Holdout-scored evals on customer-supplied examples (Tune) and eval-gated live routing receipts. A model only appears as cheapest-clearing when it empirically clears the bar at that quality threshold.

Is this page showing live measured data?

Not yet for this period. The table below is a clearly labeled methodology preview with ILLUSTRATIVE rows. When the live API publishes a period with rows, this page switches to measured aggregates automatically.

How should I cite the index?

Use: “o10 Quality-Cost Frontier Index, September 2026”. Source page: https://www.o10.io/research/quality-cost-frontier-index. Machine-readable distribution: https://app.o10.io/api/frontier-index.

o10Set the envelope. o10 holds it.

See what you're overpaying.

Paste a week of traffic. Get the number that books the audit.

See what you're overpaying
verified savings methodology · State of Inference Spend 2026