Llama 3.2 1b costs $0.100/1M input tokens and $0.100/1M output tokens as of 2026-06-27. Openrouter publishes this model with a 128K tokens context window. o10 lists Llama 3.2 1b on AWS Bedrock, Together AI, and Fireworks AI at routing band C1.

Last updated: 2026-06-27. Full catalog

Llama 3.2 1b — specs, hosts, and routing

AnswersSelf-contained

What is Llama 3.2 1b?

Llama 3.2 1b is a language model in the o10 fulfillment catalog (slug: llama-3.2-1b). Published input pricing starts at $0.100/1M tokens. Context window: 128K tokens.

Catalog snapshot 2026-06-27.

How much does Llama 3.2 1b cost per million tokens?

Llama 3.2 1b costs $0.100/1M input tokens and $0.100/1M output tokens in the gateway catalog snapshot (2026-06-27). Endpoint pricing may differ by fulfillment host.

Sourced from gateway catalog snapshot.

Where can you run Llama 3.2 1b?

Llama 3.2 1b is configured on AWS Bedrock, Together AI, and Fireworks AI. Each host maps to an upstream model ID listed in the fulfillment routes table. This model is routable through o10 eval-gated routing.

What is Llama 3.2 1b's context window?

Llama 3.2 1b supports a 128K tokens context window in the gateway catalog snapshot. Context limits may vary by fulfillment host endpoint.

What o10 routing band is Llama 3.2 1b?

Llama 3.2 1b is classified at o10 routing band C1. Bands group models by cost tier and eval profile: C0 (economy), C1 (standard), C2 (capable), C3/C4 (frontier). Production routing is enabled for this slug.

Who makes Llama 3.2 1b?

Llama 3.2 1b is attributed to source provider openrouter in the o10 catalog. Fulfillment hosts (AWS Bedrock, Together AI, and Fireworks AI) supply API access; list pricing may differ by venue.

01Deep dive

Specifications

Llama 3.2 1b in the o10 unified catalog.

Llama 3.2 1b specifications
FieldValue
o10 slug`llama-3.2-1b`
Typetext
BandC1
RoutableYes
Source provideropenrouter
Primary hostAWS Bedrock
Gateway ID`meta/llama-3.2-1b`
Context window128,000 tokens
Released2024-09-18
02Deep dive

Token pricing

Gateway snapshot pricing for Llama 3.2 1b.

Llama 3.2 1b token pricing
UnitPrice
Input$0.1/1M tokens
Output$0.1/1M tokens
03Deep dive

Fulfillment hosts

Upstream venues configured for Llama 3.2 1b.

Llama 3.2 1b fulfillment routes
HostUpstream model ID
metabedrock/us.meta.llama3-2-1b-instruct-v1:0
metatogether_ai/llama-3.2-1b
metafireworks_ai/llama-3.2-1b
04Deep dive

Gateway provider endpoints

Endpoint-level pricing and latency from the gateway catalog.

Llama 3.2 1b gateway endpoints
ProviderInput ($/1M)Latency p50Uptime 1d
bedrock$0.1218ms10000.0%
05Deep dive

Routing with o10

Eval-gated model selection for this slug.

o10 slug: llama-3.2-1b. Band: C1. Routable: yes.

See [/models/routing](/models/routing) for smart aliases (o10/auto, o10/economy) and [/kyi](/kyi) for governance.

SourceMethodology

Fulfillment catalog 2026-06-27. Gateway pricing joined where slug match exists.

FAQFrequently asked questions

Common questions

How much does Llama 3.2 1b cost?

Llama 3.2 1b costs $0.100/1M input tokens and $0.100/1M output tokens in the gateway snapshot (2026-06-27). Verify against your provider's published pricing.

What is Llama 3.2 1b's context window?

Llama 3.2 1b supports 128K tokens in the gateway catalog snapshot. Host-specific endpoints may differ.

Which hosts serve Llama 3.2 1b?

Llama 3.2 1b (llama-3.2-1b) is configured on: AWS Bedrock, Together AI, and Fireworks AI. Upstream model IDs are listed in the fulfillment routes section.

Is Llama 3.2 1b routable through o10?

Yes. Llama 3.2 1b is routable at band C1. Send model: "llama-3.2-1b" or use an o10 smart routing alias.

How do I call Llama 3.2 1b via o10?

Use o10 slug "llama-3.2-1b" or gateway ID "meta/llama-3.2-1b" depending on your integration. Gateway-compatible endpoints accept OpenAI-style /v1 requests.

How does o10 route this model?

o10 routes Llama 3.2 1b only when per-use-case evals clear at your quality floor. Band C1 indicates cost tier. See /kyi for governance and /models/routing for smart aliases.