price tracker · updated 2026-08-05
Cheapest Llama 3.3 70B API
llm · per 1M output tokens
Cheapest tracked price: $0.32/1M tokens from openrouter.
Llama 3.3 70B output tokens start at $0.32 per 1M on OpenRouter, while Cerebras and SambaNova both charge $1.20 for the same model.
Across the 6 endpoints we track for Llama 3.3 70B, posted prices span $0.32 at openrouter to $1.20 at cerebras. That is a 3.8x gap for the same model, so provider choice matters more than most teams expect. Rates below are normalized to llm · per 1M output tokens and re-checked weekly.
Llama 3.3 70B
llm · per 1M output tokens
save up to 73%
Cheapest verified endpoint: openrouter at $0.32 (llm · per 1M output tokens). Prices normalized per unit, ex. egress.
How we track this
We re-pull every known provider that serves Llama 3.3 70B weekly, normalize to a single unit (llm · per 1M output tokens), and pin the price the moment a provider posts it. We don't average across stale snapshots. Switching providers is usually a one-line base-URL change. Spot a stale price? tell us.
Frequently asked questions
What is the cheapest Llama 3.3 70B API?
openrouter currently has the lowest price for Llama 3.3 70B at $0.32 (llm · per 1M output tokens), based on our 2026-08-05 check across 6 providers.
How much does the Llama 3.3 70B API cost?
Prices for Llama 3.3 70B range from $0.32 at openrouter up to $1.20 at cerebras, measured as llm · per 1M output tokens. Same model, different bill: picking the right provider saves up to 73%.
Which providers offer Llama 3.3 70B?
We currently track 6 endpoints for Llama 3.3 70B: openrouter, groq, fireworks, together.ai, cerebras, sambanova. The table above ranks them by effective price, cheapest first.
How often are these prices updated?
We re-check every provider weekly and pin the exact posted price rather than averaging stale snapshots. This table was last refreshed on 2026-08-05.
Cut the bill further
Cheapest Inference 2026: 88 Models From $0.04 per 1M
Cost · 6 min read
The Cheapest Way to Run LLMs in 2026
Cost · 7 min read
Prompt Caching and Batch APIs: Cut Your LLM Bill in Half
Cost · 6 min read
Similar models we track
Cheapest DeepSeek-V3 API
llm · per 1M output tokens · save up to 41%
Cheapest DeepSeek-R1 API
llm · per 1M output tokens · save up to 69%
Cheapest Llama 4 Maverick API
llm · per 1M output tokens · save up to 29%
Cheapest Llama 4 Scout API
llm · per 1M output tokens · save up to 49%
Cheapest Llama 3.1 405B API
llm · per 1M output tokens · save up to 13%
Cheapest Llama 3.1 8B API
llm · per 1M output tokens · save up to 72%
Cheapest Qwen2.5 72B API
llm · per 1M output tokens · save up to 67%
Cheapest Qwen3 235B A22B API
llm · per 1M output tokens · save up to 86%
Cheapest Mistral Nemo API
llm · per 1M output tokens · save up to 76%
Cheapest Mistral Small 3 API
llm · per 1M output tokens · save up to 73%
Cheapest Kimi K2 API
llm · per 1M output tokens · save up to 56%
Cheapest GLM-4.6 API
llm · per 1M output tokens
Browse all 89 model price trackers →
Building on AI? Don't pay full price.
Perkstack also tracks 230+ verified AI tool credits and startup grants, each with its verified conditions and apply link.