# GPT OSS 120B API prices

GPT OSS 120B (gpt-oss) — 131k context, 131k max output.
74 metered per-token offers from 61 provider(s), in USD per 1M tokens.
Prices as of 2026-10-04.

| Provider | Input $/1M | Output $/1M | Cache read $/1M | Context |
| --- | --- | --- | --- | --- |
| DevPass (LLM Gateway) | $0.032 | $0.14 | $0.032 | 131k |
| Kilo Gateway | $0.03 | $0.17 | $0.03 | 131k |
| OrcaRouter | $0.03 | $0.17 | — | 131k |
| CoreWeave | $0.03 | $0.17 | $0.03 | 131k |
| Helicone | $0.04 | $0.16 | — | 131k |
| Deep Infra | $0.037 | $0.17 | — | 131k |
| Eden AI | $0.037 | $0.17 | — | 131k |
| OpenRouter | $0.037 | $0.17 | — | 131k |
| Crusoe | $0.05 | $0.2 | $0.05 | 131k |
| Merge Gateway | $0.05 | $0.25 | — | 128k |
| NovitaAI | $0.05 | $0.25 | — | 131k |
| Synthetic | $0.1 | $0.1 | $0.1 | 131k |
| DInference | $0.0675 | $0.27 | — | 131k |
| Databricks | $0.072 | $0.28 | — | 131k |
| Venice AI | $0.07 | $0.3 | — | 128k |
| IO.NET | $0.04 | $0.4 | $0.02 | 131k |
| SiliconFlow | $0.05 | $0.45 | — | 131k |
| Vertex | $0.09 | $0.36 | — | 131k |
| Abacus | $0.08 | $0.44 | — | 128k |
| Cortecs | $0.089 | $0.446 | $0.01 | 131k |
| OpenReason | $0.1055 | $0.422 | — | 131k |
| Eden AI | $0.09 | $0.47 | — | 131k |
| OVHcloud AI Endpoints | $0.09 | $0.47 | — | 131k |
| submodel | $0.1 | $0.5 | — | 131k |
| Vercel AI Gateway | $0.1 | $0.5 | $0.1 | 131k |
| AKI.IO | $0.15 | $0.55 | — | 128k |
| DigitalOcean | $0.1 | $0.7 | $0.02 | 128k |
| ai& | $0.15 | $0.6 | $0.08 | 131k |
| Amazon Bedrock | $0.15 | $0.6 | — | 131k |
| Amazon Bedrock | $0.15 | $0.6 | — | 131k |
| Eden AI | $0.15 | $0.6 | $0.015 | 131k |
| Eden AI | $0.15 | $0.6 | $0.014 | 131k |
| Eden AI | $0.15 | $0.6 | $0.075 | 131k |
| Eden AI | $0.15 | $0.6 | $0.15 | 131k |
| Eden AI | $0.15 | $0.6 | — | 131k |
| FastRouter | $0.15 | $0.6 | — | 131k |
| Fireworks AI | $0.15 | $0.6 | $0.015 | 131k |
| FrogBot | $0.15 | $0.6 | — | 131k |
| Groq | $0.15 | $0.6 | $0.075 | 131k |
| Impossibl | $0.15 | $0.6 | $0.015 | 131k |
| Impossibl | $0.15 | $0.6 | $0.075 | 131k |
| LLM Gateway | $0.15 | $0.6 | — | 131k |
| LLM Gateway | $0.15 | $0.6 | — | 131k |
| Nebius Token Factory | $0.15 | $0.6 | $0.015 | 131k |
| Neon | $0.15 | $0.6 | — | 131k |
| OCI Generative AI | $0.15 | $0.6 | — | 128k |
| Ollama Cloud | $0.15 | $0.6 | $0.014 | 131k |
| Pioneer | $0.15 | $0.6 | $0.015 | 131k |
| Scaleway | $0.15 | $0.6 | — | 128k |
| Tempr Gateway | $0.15 | $0.6 | $0.075 | 131k |
| Tinfoil | $0.15 | $0.6 | — | 131k |
| Together AI | $0.15 | $0.6 | — | 131k |
| Eden AI | $0.15 | $0.6 | $0.015 | 131k |
| SCX.ai | $0.17 | $0.55 | — | 131k |
| watsonx.ai | $0.159 | $0.636 | — | 131k |
| Eden AI | $0.1684 | $0.6735 | — | 128k |
| LLM Gateway | $0.15 | $0.75 | — | 131k |
| Charm Hyper | $0.178 | $0.68 | $0.089 | 131k |
| Hugging Face | $0.25 | $0.69 | — | 131k |
| Privatemode AI | $0.2311 | $0.7511 | $0.0462 | 128k |
| GreenPT | $0.228 | $0.798 | — | 131k |
| evroc | $0.23 | $0.92 | — | 66k |
| Cerebras | $0.35 | $0.75 | — | 131k |
| Cloudflare Workers AI | $0.35 | $0.75 | — | 128k |
| Eden AI | $0.35 | $0.75 | $0.35 | 131k |
| Eden AI | $0.35 | $0.75 | — | 128k |
| Impossibl | $0.35 | $0.75 | — | 131k |
| LLM Gateway | $0.35 | $0.75 | — | 131k |
| NanoGPT | $0.35 | $0.75 | — | 128k |
| Tempr Gateway | $0.35 | $0.75 | — | 131k |
| STACKIT | $0.53 | $0.76 | — | 131k |
| Regolo AI | $1 | $4.2 | — | 128k |
| Opper | $1.1622 | $4.8812 | — | 128k |
| CloudFerro Sherlock | $2.92 | $2.92 | — | 131k |

## Summary

- Cheapest input: $0.03 per 1M tokens (Kilo Gateway)
- Cheapest output: $0.1 per 1M tokens (Synthetic)
- First-party: not listed separately
- Free offers: Kenari, Pendra, QVAC
- Subscription plans (not per-token): none
- Released: 2025-08-05
- Inputs: image, text

## FAQ

### What is the cheapest GPT OSS 120B API?

As of Oct 4, 2026, Synthetic has the lowest GPT OSS 120B output price at $0.1 per 1M tokens, and Kilo Gateway has the lowest input price at $0.03 per 1M tokens. Prices are metered per-token rates in USD; free tiers and subscription plans are excluded.

### How many providers offer GPT OSS 120B?

61 providers list GPT OSS 120B on Sovyron; 74 of them sell it at a metered per-token price.

### Is GPT OSS 120B free?

3 provider(s) list a free-tier offer: Kenari, Pendra, QVAC. Free tiers usually have rate limits.

### What is the context window of GPT OSS 120B?

GPT OSS 120B supports a 131k-token context window and up to 131k output tokens.

HTML page: https://sovyron.com/models/gpt-oss-120b/
Full dataset: https://sovyron.com/data/catalog.json
