Llama 4 Scout costs $0.170/1M input tokens and $0.660/1M output tokens as of 2026-06-27. Openrouter publishes this model with a 128K tokens context window. o10 lists Llama 4 Scout on AWS Bedrock, Together AI, and Fireworks AI at routing band C3. API slug llama-4-scout. Verify list prices against the venue you call.

Last updated: 2026-06-27. Full catalog. Reference prices are dated snapshots. Compare total cost and measured task quality before choosing a route.

Llama 4 Scout. Price, context, and routing

AnswersSelf-contained

What is Llama 4 Scout?

Llama 4 Scout is a chat model from openrouter (slug: llama-4-scout). Pick Llama 4 Scout when you want Meta Llama on a fulfillment host (or your own GPUs) instead of a closed lab SKU. Published input pricing starts at $0.170/1M tokens. Context window: 128K tokens.

Catalog snapshot 2026-06-27.

When should you pick Llama 4 Scout?

Pick Llama 4 Scout when you want Meta Llama on a fulfillment host (or your own GPUs) instead of a closed lab SKU. Llama 4 Scout is a Llama instruct checkpoint. Smaller Llama SKUs cut cost; larger ones raise quality. Prove the swap in shadow.

How does Llama 4 Scout differ from sibling models?

Llama 4 Scout is a Llama instruct checkpoint. Smaller Llama SKUs cut cost; larger ones raise quality. Prove the swap in shadow.

How much does Llama 4 Scout cost per million tokens?

Llama 4 Scout costs $0.170/1M input tokens and $0.660/1M output tokens in the gateway catalog snapshot (2026-06-27). Endpoint pricing may differ by fulfillment host.

Sourced from gateway catalog snapshot.

Where can you run Llama 4 Scout?

Llama 4 Scout is configured on AWS Bedrock, Together AI, and Fireworks AI. Known host IDs: meta: bedrock/us.meta.llama4-scout-17b-instruct-v1:0; meta: together_ai/llama-4-scout; meta: fireworks_ai/llama-4-scout. This model is routable through o10 eval-gated routing.

What is Llama 4 Scout's context window?

Llama 4 Scout supports a 128K tokens context window in the gateway catalog snapshot. Context limits may vary by fulfillment host endpoint.

What o10 routing band is Llama 4 Scout?

Llama 4 Scout is classified at o10 routing band C3. Bands group models by cost tier and eval profile: C0 (economy), C1 (standard), C2 (capable), C3/C4 (frontier). Production routing is enabled for this slug.

Who makes Llama 4 Scout?

Llama 4 Scout is attributed to source provider openrouter in the o10 catalog. Fulfillment hosts (AWS Bedrock, Together AI, and Fireworks AI) supply API access; list pricing may differ by venue.

01Deep dive

Specifications

Llama 4 Scout in the o10 unified catalog.

Llama 4 Scout specifications
FieldValue
o10 slug`llama-4-scout`
Typetext
BandC3
RoutableYes
Source provideropenrouter
Primary hostAWS Bedrock
Gateway ID`meta/llama-4-scout`
Context window128,000 tokens
Released2025-04-05
02Deep dive

Token pricing

Gateway snapshot pricing for Llama 4 Scout.

Llama 4 Scout token pricing
UnitPrice
Input$0.17/1M tokens
Output$0.66/1M tokens
03Deep dive

Fulfillment hosts

Upstream venues configured for Llama 4 Scout.

Llama 4 Scout fulfillment routes
HostUpstream model ID
metabedrock/us.meta.llama4-scout-17b-instruct-v1:0
metatogether_ai/llama-4-scout
metafireworks_ai/llama-4-scout
04Deep dive

Gateway provider endpoints

Endpoint-level pricing and latency from the gateway catalog.

Llama 4 Scout gateway endpoints
ProviderInput ($/1M)Latency p50Uptime 1d
bedrock$0.17231.5ms10000.0%
deepinfra$0.1204ms10000.0%
groq$0.11168.5ms10000.0%
05Deep dive

Routing with o10

Eval-gated model selection for this slug.

o10 slug: llama-4-scout. Band: C3. Routable: yes.

See [/models/routing](/models/routing) for smart aliases (o10/auto, o10/frontier, o10/squad) and [/kyi](/kyi) for governance.

SourceMethodology

Fulfillment catalog 2026-06-27. Gateway pricing joined where slug match exists.

FAQFrequently asked questions

Common questions

How much does Llama 4 Scout cost?

Llama 4 Scout costs $0.170/1M input tokens and $0.660/1M output tokens in the gateway snapshot (2026-06-27). Verify against your provider's published pricing.

What is Llama 4 Scout's context window?

Llama 4 Scout supports 128K tokens in the gateway catalog snapshot. Host-specific endpoints may differ.

Which hosts serve Llama 4 Scout?

Llama 4 Scout (llama-4-scout) is configured on: AWS Bedrock, Together AI, and Fireworks AI. Upstream model IDs are listed in the fulfillment routes section.

Is Llama 4 Scout routable through o10?

Yes. Llama 4 Scout is routable at band C3. Send model: "llama-4-scout" or use an o10 smart routing alias.

How do I call Llama 4 Scout via o10?

Use o10 slug "llama-4-scout" or gateway ID "meta/llama-4-scout" depending on your integration. Gateway-compatible endpoints accept OpenAI-style /v1 requests.

How does o10 route this model?

o10 routes Llama 4 Scout only when per-use-case evals clear at your quality floor. Band C3 indicates cost tier. See /kyi for governance and /models/routing for smart aliases.