Llama 3.3 70B Instruct costs $0.900/1M input tokens and $0.900/1M output tokens as of 2026-06-27. Meta publishes this model with a 128K tokens context window. o10 lists Llama 3.3 70B Instruct on NVIDIA NIM and SambaNova at routing band C2. API slug meta-llama-3-3-70b-instruct. Verify list prices against the venue you call.

Last updated: 2026-06-27. Full catalog. Reference prices are dated snapshots. Compare total cost and measured task quality before choosing a route.

Llama 3.3 70B Instruct pricing. Input/output per 1M tokens

AnswersSelf-contained

How much does Llama 3.3 70B Instruct cost per million tokens?

Llama 3.3 70B Instruct costs $0.900/1M input tokens and $0.900/1M output tokens in the gateway catalog snapshot (2026-06-27). Endpoint pricing may differ by fulfillment host.

Sourced from gateway catalog snapshot.

Where can you run Llama 3.3 70B Instruct?

Llama 3.3 70B Instruct is configured on NVIDIA NIM and SambaNova. Known host IDs: nvidia_nim: nvidia_nim/meta/llama-3.3-70b-instruct; nvidia_nim: sambanova/Meta-Llama-3.3-70B-Instruct. This model is routable through o10 eval-gated routing.

What is Llama 3.3 70B Instruct's context window?

Llama 3.3 70B Instruct supports a 128K tokens context window in the gateway catalog snapshot. Context limits may vary by fulfillment host endpoint.

01Deep dive

Token pricing

Gateway snapshot pricing for Llama 3.3 70B Instruct.

Llama 3.3 70B Instruct token pricing
UnitPrice
Input$0.9/1M tokens
Output$0.9/1M tokens
02Deep dive

Fulfillment hosts

Upstream venues configured for Llama 3.3 70B Instruct.

Llama 3.3 70B Instruct fulfillment routes
HostUpstream model ID
nvidia_nimnvidia_nim/meta/llama-3.3-70b-instruct
nvidia_nimsambanova/Meta-Llama-3.3-70B-Instruct
03Deep dive

Routing with o10

Eval-gated model selection for this slug.

o10 slug: meta-llama-3-3-70b-instruct. Band: C2. Routable: yes.

See [/models/routing](/models/routing) for smart aliases (o10/auto, o10/frontier, o10/squad) and [/kyi](/kyi) for governance.

SourceMethodology

Gateway catalog snapshot 2026-06-27. Fulfillment catalog synced from o10 control plane. Verify list prices against meta and host providers.

FAQFrequently asked questions

Common questions

How much does Llama 3.3 70B Instruct cost?

Llama 3.3 70B Instruct costs $0.900/1M input tokens and $0.900/1M output tokens in the gateway snapshot (2026-06-27). Verify against your provider's published pricing.

What is Llama 3.3 70B Instruct's context window?

Llama 3.3 70B Instruct supports 128K tokens in the gateway catalog snapshot. Host-specific endpoints may differ.

Which hosts serve Llama 3.3 70B Instruct?

Llama 3.3 70B Instruct (meta-llama-3-3-70b-instruct) is configured on: NVIDIA NIM and SambaNova. Upstream model IDs are listed in the fulfillment routes section.

Is Llama 3.3 70B Instruct routable through o10?

Yes. Llama 3.3 70B Instruct is routable at band C2. Send model: "meta-llama-3-3-70b-instruct" or use an o10 smart routing alias.

How do I call Llama 3.3 70B Instruct via o10?

Use o10 slug "meta-llama-3-3-70b-instruct" in the model field. See /models/routing for smart aliases (o10/auto, o10/frontier, o10/squad).

How does o10 route this model?

o10 routes Llama 3.3 70B Instruct only when per-use-case evals clear at your quality floor. Band C2 indicates cost tier. See /kyi for governance and /models/routing for smart aliases.

What is the API model ID for Llama 3.3 70B Instruct?

The o10 slug is "meta-llama-3-3-70b-instruct". Host IDs: nvidia_nim: nvidia_nim/meta/llama-3.3-70b-instruct; nvidia_nim: sambanova/Meta-Llama-3.3-70B-Instruct. Pricing snapshot date: 2026-06-27.

Who hosts Llama 3.3 70B Instruct?

Llama 3.3 70B Instruct is hosted as nvidia_nim: nvidia_nim/meta/llama-3.3-70b-instruct; nvidia_nim: sambanova/Meta-Llama-3.3-70B-Instruct. o10 maps those upstream IDs to slug "meta-llama-3-3-70b-instruct".