Free prompt optimization. Send a prompt and 5–20 real examples. o10 finds a cheaper model and a tuned prompt that still clear your quality bar, and hands you a before/after receipt.
Prompt-optimization tools often charge per successful run and bill the optimization's LLM usage on top.
Recommendation-layer tools often charge for routing and/or bill optimization compute separately (Not Diamond's public pricing lists ~$0.05 per million tokens routed as of 2026-07; a fixed ~$20-per-success optimization fee is not listed on their public pricing page, needs operator confirmation before citing). o10 Tune bundles the optimization loop on low-cost routed supply, so successful tunes are free and failures never charge.
o10 runs the loop on its own low-cost routed supply, so tuning is free. Failures never charge either.
04Deep dive
Where it works
Same Tune loop across console, API, Slack, and DeepShell.
Console: https://app.o10.io/tune. API: POST /v1/optimizations. Slack: @o10 tune. DeepShell: "make this task cheaper". See /deepshell.
Console: app.o10.io/tune
API: POST /v1/optimizations
Slack: @o10 tune
DeepShell: make this task cheaper
How-toOperational steps
Get started
01
Start free
Sign up at https://app.o10.io/signup. Free route key, no card.
02
Open Tune
Go to https://app.o10.io/tune with a prompt and 5–20 real examples.
03
Ship the receipt
Adopt the tuned prompt + recommended model; keep the before/after receipt.
Free prompt optimization. Submit a prompt and 5–20 real examples; o10 searches cheaper models and prompt variants and returns a tuned prompt + recommended model with a before/after receipt. Quality is measured on a held-out split of your examples. Successful tunes are free; failures never charge.
Is Tune really free?
Successful tunes are free; failures never charge. o10 runs the optimization loop on its own low-cost routed supply, so you are not billed separately for the optimization's LLM usage.
How does quality stay honest?
Quality is measured on a held-out split of your own examples. Results are never overstated. You get a before/after receipt: quality, cost per 1K calls, and savings %.
Where can I run Tune?
Console (https://app.o10.io/tune), API (POST /v1/optimizations), Slack (@o10 tune), and DeepShell ("make this task cheaper").
How is Tune different from Not Diamond?
Not Diamond is a recommendation layer. It tells you which model to call; you typically pay your own provider bills, and public pricing lists ~$0.05 per million tokens routed (verify at notdiamond.ai/pricing). o10 is an in-path gateway that serves tokens, enforces the quality floor and budget, and issues savings receipts. Recommendation-layer tools often charge for routing and/or bill optimization compute separately (Not Diamond's public pricing lists ~$0.05 per million tokens routed as of 2026-07; a fixed ~$20-per-success optimization fee is not listed on their public pricing page, needs operator confirmation before citing). o10 Tune bundles the optimization loop on low-cost routed supply, so successful tunes are free and failures never charge.
How do I start Tune?
Start free at https://app.o10.io/signup, then open https://app.o10.io/tune or POST POST /v1/optimizations with a prompt and 5–20 examples.
o10Set the envelope. o10 holds it.
See what you're overpaying.
Paste a week of traffic. Get the number that books the audit.