Inference providers · open weights · Meta

Llama 3.3 70B Instruct

Open-weights model published by Meta; the API prices below are set by the 2 hosts that serve it, not by Meta. There is no maker's list price to compare with: Meta publishes the weights, not a per-token rate. Each figure was read on that host's own pricing page and links to it, with the date it was checked.

What each host charges

2 hosts, from $0.10 (DeepInfra) to $1.04 (Together AI) per 1M input tokens: a 10× spread. Above the 1.15× line this site draws, so the choice of host is a price decision.

Host Input Output Cached input Context served Checked Source
DeepInfra $0.10 $0.32 — 131K tokens September 30, 2026 deepinfra.com
Together AI $1.04 $1.04 — 131K tokens October 1, 2026 together.ai

DeepInfra charges $0.10 per 1M input tokens and Together AI $1.04 for the same weights; on output, $0.32 against $1.04.

Hosts that no longer price it

Offerings this site tracked and stopped listing, with the date and the reason. They are kept because nobody else records when a host drops a model.

On OpenRouter

On September 15, 2026 OpenRouter listed 11 endpoints for Llama 3.3 70B Instruct, 9 of them healthy, from $0.10 (DeepInfra) to $1.04 per 1M input tokens. The figure OpenRouter's catalog shows for the model, $0.10 / $0.32 (DeepInfra), is the cheapest of those 11. None of them is Meta: all are resellers, which is why this site shows no OpenRouter price for the model. DeepInfra and Together AI, priced above from their own pages, are among them. Groq, which this site stopped listing on September 6, 2026 because its own pricing page carries no price for the model, serves it once through OpenRouter; that is a reseller listing, not a published price, so the row stays out.

Specifications

Weights published by Meta
Context length (maker) 131K tokens
Context served by hosts 131K tokens
API model ids DeepInfra: meta-llama/Llama-3.3-70B-Instruct-Turbo; Together AI: meta-llama/Llama-3.3-70B-Instruct-Turbo
Hosts with a verified price 2

Host prices verified between September 30, 2026 and October 1, 2026, each on the host's own pricing page linked in the table above, never from a reseller catalog or a third-party table. Maker and context length from Meta's model card, checked September 15, 2026 (Context length 128k as stated on the model card). How the data is checked.