Source page: https://www.o10.io/models/meta-llama-3-3-70b-instruct

Plain-text reference for reading, copying, and citation. Examples are illustrative; use your own credentials for API requests.

Model data: https://www.o10.io/api/models.json
Content index: https://www.o10.io/llms.txt
Prices are dated references, not live quotes. Check the current endpoint rate before use.

# Llama 3.3 70B Instruct. Price, context, and routing

Llama 3.3 70B Instruct costs $0.900/1M input tokens and $0.900/1M output tokens as of 2026-06-27. Meta publishes this model with a 128K tokens context window. o10 lists Llama 3.3 70B Instruct on NVIDIA NIM and SambaNova at routing band C2. API slug meta-llama-3-3-70b-instruct. Verify list prices against the venue you call.

## Key takeaways

### What is Llama 3.3 70B Instruct?

Llama 3.3 70B Instruct is a chat model from meta (slug: meta-llama-3-3-70b-instruct). Pick Llama 3.3 70B Instruct when you want Meta Llama on a fulfillment host (or your own GPUs) instead of a closed lab SKU. Published input pricing starts at $0.900/1M tokens. Context window: 128K tokens.

*Catalog snapshot 2026-06-27.*

### When should you pick Llama 3.3 70B Instruct?

Pick Llama 3.3 70B Instruct when you want Meta Llama on a fulfillment host (or your own GPUs) instead of a closed lab SKU. Llama 3.3 70B Instruct is a Llama instruct checkpoint. Smaller Llama SKUs cut cost; larger ones raise quality. Prove the swap in shadow.

### How does Llama 3.3 70B Instruct differ from sibling models?

Llama 3.3 70B Instruct is a Llama instruct checkpoint. Smaller Llama SKUs cut cost; larger ones raise quality. Prove the swap in shadow.

### How much does Llama 3.3 70B Instruct cost per million tokens?

Llama 3.3 70B Instruct costs $0.900/1M input tokens and $0.900/1M output tokens in the gateway catalog snapshot (2026-06-27). Endpoint pricing may differ by fulfillment host.

*Sourced from gateway catalog snapshot.*

### Where can you run Llama 3.3 70B Instruct?

Llama 3.3 70B Instruct is configured on NVIDIA NIM and SambaNova. Known host IDs: nvidia_nim: nvidia_nim/meta/llama-3.3-70b-instruct; nvidia_nim: sambanova/Meta-Llama-3.3-70B-Instruct. This model is routable through o10 eval-gated routing.

### What is Llama 3.3 70B Instruct's context window?

Llama 3.3 70B Instruct supports a 128K tokens context window in the gateway catalog snapshot. Context limits may vary by fulfillment host endpoint.

### What o10 routing band is Llama 3.3 70B Instruct?

Llama 3.3 70B Instruct is classified at o10 routing band C2. Bands group models by cost tier and eval profile: C0 (economy), C1 (standard), C2 (capable), C3/C4 (frontier). Production routing is enabled for this slug.

### Who makes Llama 3.3 70B Instruct?

Llama 3.3 70B Instruct is attributed to source provider meta in the o10 catalog. Fulfillment hosts (NVIDIA NIM and SambaNova) supply API access; list pricing may differ by venue.

## Citation
Llama 3.3 70B Instruct costs $0.900/1M input tokens and $0.900/1M output tokens as of 2026-06-27. Meta publishes this model with a 128K tokens context window. o10 lists Llama 3.3 70B Instruct on NVIDIA NIM and SambaNova at routing band C2. API slug meta-llama-3-3-70b-instruct. Verify list prices against the venue you call.
## Methodology

Fulfillment catalog 2026-06-27. Gateway pricing joined where slug match exists.

## FAQ

### How much does Llama 3.3 70B Instruct cost?

Llama 3.3 70B Instruct costs $0.900/1M input tokens and $0.900/1M output tokens in the gateway snapshot (2026-06-27). Verify against your provider's published pricing.

### What is Llama 3.3 70B Instruct's context window?

Llama 3.3 70B Instruct supports 128K tokens in the gateway catalog snapshot. Host-specific endpoints may differ.

### Which hosts serve Llama 3.3 70B Instruct?

Llama 3.3 70B Instruct (meta-llama-3-3-70b-instruct) is configured on: NVIDIA NIM and SambaNova. Upstream model IDs are listed in the fulfillment routes section.

### Is Llama 3.3 70B Instruct routable through o10?

Yes. Llama 3.3 70B Instruct is routable at band C2. Send model: "meta-llama-3-3-70b-instruct" or use an o10 smart routing alias.

### How do I call Llama 3.3 70B Instruct via o10?

Use o10 slug "meta-llama-3-3-70b-instruct" in the model field. See /models/routing for smart aliases (o10/auto, o10/frontier, o10/squad).

### How does o10 route this model?

o10 routes Llama 3.3 70B Instruct only when per-use-case evals clear at your quality floor. Band C2 indicates cost tier. See /kyi for governance and /models/routing for smart aliases.

### What is the API model ID for Llama 3.3 70B Instruct?

The o10 slug is "meta-llama-3-3-70b-instruct". Host IDs: nvidia_nim: nvidia_nim/meta/llama-3.3-70b-instruct; nvidia_nim: sambanova/Meta-Llama-3.3-70B-Instruct. Pricing snapshot date: 2026-06-27.

### Who hosts Llama 3.3 70B Instruct?

Llama 3.3 70B Instruct is hosted as nvidia_nim: nvidia_nim/meta/llama-3.3-70b-instruct; nvidia_nim: sambanova/Meta-Llama-3.3-70B-Instruct. o10 maps those upstream IDs to slug "meta-llama-3-3-70b-instruct".

## Related links

- [Llama 3.3 70B Instruct HTML](https://www.o10.io/models/meta-llama-3-3-70b-instruct)
- [Llama 3.3 70B Instruct pricing](https://www.o10.io/pricing/meta/meta-llama-3-3-70b-instruct)
- [Model catalog](https://www.o10.io/models)
- [llms-models.txt](https://www.o10.io/llms-models.txt)

## Source URL

https://www.o10.io/models/meta-llama-3-3-70b-instruct
