Skip to content
Sovyron

watsonx.ai

5 modelsProvider docs

As of Oct 4, 2026, watsonx.ai lists 5 supported models in the Sovyron catalog, including Granite-4.0-H-Small, GPT OSS 120B, Llama 4 Maverick 17B 128E Instruct FP8, Mistral Small 3.1 24B and Llama-3.3-70B-Instruct — granite, gpt-oss, llama and mistral-small families. Its metered rates span $0.114–$0.7526 per 1M tokens on a 3:1 input:output blend, and every price in the table is watsonx.ai's own, normalized to USD — open a model to compare it with the cheapest offer across all providers.

Supported models and pricing

5 of 5 models
watsonx.ai model prices in USD per 1M tokens, updated Oct 4, 2026
Model1M input1M output1M cache readContextReleased
Granite-4.0-H-Small$0.0636$0.265—131kOct 2, 2025
GPT OSS 120B$0.159$0.636—131kAug 5, 2025
Llama 4 Maverick 17B 128E Instruct FP8$0.371$1.484—131kApr 5, 2025
Mistral Small 3.1 24B$0.106$0.318—131kMar 17, 2025
Llama-3.3-70B-Instruct$0.7526$0.7526—131kDec 6, 2024