Skip to content
Sovyron

Together AI

18 modelsProvider docs

As of Oct 4, 2026, Together AI lists 18 supported models in the Sovyron catalog, including DeepSeek V4.1 Flash, GLM-5.3-Flash, GLM-5.3, DeepSeek V4 Pro 0813 and DeepSeek V4 Flash 0731 — deepseek-flash, glm-flash, glm, deepseek-thinking, kimi-k3 and ling families. Its metered rates span $0.0525–$6 per 1M tokens on a 3:1 input:output blend, and every price in the table is Together AI's own, normalized to USD — open a model to compare it with the cheapest offer across all providers.

Supported models and pricing

18 of 18 models
Together AI model prices in USD per 1M tokens, updated Oct 4, 2026
Model1M input1M output1M cache readContextReleased
DeepSeek V4.1 Flash$0.3$1.2$0.0061.0MSep 10, 2026
GLM-5.3-Flash$0.15$0.5$0.031.0MAug 26, 2026
GLM-5.3$1.4$4.4$0.261.0MAug 14, 2026
DeepSeek V4 Pro 0813$1.32$3.96$0.131.0MAug 12, 2026
DeepSeek V4 Flash 0731$0.14$0.28$0.031.0MJul 31, 2026
Kimi K3$3$15$0.31.0MJul 16, 2026
Inkling$1$4.05$0.17524kJul 15, 2026
GLM-5.2$1.4$4.4$0.261.0MJun 16, 2026
MiniMax-M3$0.3$1.2$0.06524kJun 12, 2026
Nemotron 3 Ultra 550B A55B$0.6$3.6$0.2512kJun 4, 2026
Qwen3.7 Max$1.25$3.75$0.1251.0MMay 21, 2026
Qwen3.6 Plus$0.5$3—1.0MApr 30, 2026
MiniMax-M2.7$0.3$1.2$0.06197kMar 18, 2026
Qwen3.5 9B$0.17$0.25—262kMar 3, 2026
LFM2 24B A2B$0.03$0.12—33kFeb 25, 2026
GPT OSS 120B$0.15$0.6—131kAug 5, 2025
Llama 3.3 70B$1.04$1.04—131kDec 6, 2024
Qwen 2.5 7B Instruct Turbo$0.3$0.3—33kSep 19, 2024