Skip to content
Sovyron

Modal

4 modelsProvider docs

As of Oct 4, 2026, Modal lists 4 supported models in the Sovyron catalog, including GLM-5.3-Flash, Qwen3.8 Max, Kimi K3 and Inkling — glm-flash, qwen, kimi-k3 and ling families. Its metered rates span $0.7125–$6 per 1M tokens on a 3:1 input:output blend, and every price in the table is Modal's own, normalized to USD — open a model to compare it with the cheapest offer across all providers.

Supported models and pricing

4 of 4 models
Modal model prices in USD per 1M tokens, updated Oct 4, 2026
Model1M input1M output1M cache readContextReleased
GLM-5.3-Flash$0.45$1.5$0.091.0MAug 26, 2026
Qwen3.8 Max$2$6$0.251.0MAug 12, 2026
Kimi K3$3$15$0.31.0MJul 16, 2026
Inkling$1.2$5$0.271.0MJul 15, 2026