Llama 3.3 70B Versatile
As of Oct 4, 2026, the cheapest Llama 3.3 70B Versatile API price is $0.59 per 1M input tokens (Abacus) and $0.79 per 1M output tokens (Helicone). Llama 3.3 70B Versatile is offered by 2 providers in the llama family, open weights.
API pricing by provider
Metered per-token prices. Cheapest input and output are highlighted; † marks a context-tier surcharge. How prices are ranked.
Llama 3.3 70B Versatile pricing FAQ
What is the cheapest Llama 3.3 70B Versatile API?
As of Oct 4, 2026, Helicone has the lowest Llama 3.3 70B Versatile output price at $0.79 per 1M tokens, and Abacus has the lowest input price at $0.59 per 1M tokens. Prices are metered per-token rates in USD; free tiers and subscription plans are excluded.
How many providers offer Llama 3.3 70B Versatile?
2 providers list Llama 3.3 70B Versatile on Sovyron; 2 of them sell it at a metered per-token price.
Is Llama 3.3 70B Versatile free?
No provider in the Sovyron catalog lists a free tier for Llama 3.3 70B Versatile.
What is the context window of Llama 3.3 70B Versatile?
Llama 3.3 70B Versatile supports a 131k-token context window and up to 33k output tokens.
Token cost calculator
What a month costs at one provider's published rates. Enter the tokens you expect to send and receive; cached tokens use the cached-input price where the provider lists one, and count as ordinary input where it does not.
Estimated
—