Llama 3.2 11b costs $0.160/1M input tokens and $0.160/1M output tokens as of 2026-06-27. Openrouter publishes this model with a 128K tokens context window. o10 lists Llama 3.2 11b on AWS Bedrock, Together AI, and Fireworks AI at routing band C2.
Llama 3.2 11b pricing — input/output per 1M tokens
AnswersSelf-contained
How much does Llama 3.2 11b cost per million tokens?
Llama 3.2 11b costs $0.160/1M input tokens and $0.160/1M output tokens in the gateway catalog snapshot (2026-06-27). Endpoint pricing may differ by fulfillment host.
Sourced from gateway catalog snapshot.
Where can you run Llama 3.2 11b?
Llama 3.2 11b is configured on AWS Bedrock, Together AI, and Fireworks AI. Each host maps to an upstream model ID listed in the fulfillment routes table. This model is routable through o10 eval-gated routing.
What is Llama 3.2 11b's context window?
Llama 3.2 11b supports a 128K tokens context window in the gateway catalog snapshot. Context limits may vary by fulfillment host endpoint.
01Deep dive
Token pricing
Gateway snapshot pricing for Llama 3.2 11b.
Llama 3.2 11b token pricing
Unit
Price
Input
$0.16/1M tokens
Output
$0.16/1M tokens
02Deep dive
Fulfillment hosts
Upstream venues configured for Llama 3.2 11b.
Llama 3.2 11b fulfillment routes
Host
Upstream model ID
meta
bedrock/us.meta.llama3-2-11b-instruct-v1:0
meta
together_ai/llama-3.2-11b
meta
fireworks_ai/llama-3.2-11b
03Deep dive
Routing with o10
Eval-gated model selection for this slug.
o10 slug: llama-3.2-11b. Band: C2. Routable: yes.
See [/models/routing](/models/routing) for smart aliases (o10/auto, o10/frontier, o10/squad) and [/kyi](/kyi) for governance.
SourceMethodology
Gateway catalog snapshot 2026-06-27. Fulfillment catalog synced from o10 control plane. Verify list prices against openrouter and host providers.
Llama 3.2 11b costs $0.160/1M input tokens and $0.160/1M output tokens in the gateway snapshot (2026-06-27). Verify against your provider's published pricing.
What is Llama 3.2 11b's context window?
Llama 3.2 11b supports 128K tokens in the gateway catalog snapshot. Host-specific endpoints may differ.
Which hosts serve Llama 3.2 11b?
Llama 3.2 11b (llama-3.2-11b) is configured on: AWS Bedrock, Together AI, and Fireworks AI. Upstream model IDs are listed in the fulfillment routes section.
Is Llama 3.2 11b routable through o10?
Yes. Llama 3.2 11b is routable at band C2. Send model: "llama-3.2-11b" or use an o10 smart routing alias.
How do I call Llama 3.2 11b via o10?
Use o10 slug "llama-3.2-11b" or gateway ID "meta/llama-3.2-11b" depending on your integration. Gateway-compatible endpoints accept OpenAI-style /v1 requests.
How does o10 route this model?
o10 routes Llama 3.2 11b only when per-use-case evals clear at your quality floor. Band C2 indicates cost tier. See /kyi for governance and /models/routing for smart aliases.