Skip to content
Sovyron

Privatemode AI

9 modelsProvider docs

As of Oct 4, 2026, Privatemode AI lists 9 supported models in the Sovyron catalog, including GLM Flash Latest, GLM-5.3-Flash, GLM Latest, GLM-5.3 and DeepSeek OCR 2 — glm-flash, glm, gpt-oss, voxtral, qwen and whisper families. Its metered rates span $0.0035–$3.0014 per 1M tokens on a 3:1 input:output blend, and every price in the table is Privatemode AI's own, normalized to USD — open a model to compare it with the cheapest offer across all providers.

Supported models and pricing

9 of 9 models
Privatemode AI model prices in USD per 1M tokens, updated Oct 4, 2026
Model1M input1M output1M cache readContextReleased
GLM Flash Latest$0.2311$0.7511$0.05781.0MAug 26, 2026
GLM-5.3-Flash$0.2311$0.7511$0.05781.0MAug 26, 2026
GLM Latest$1.791$6.6326$0.17331.0MAug 14, 2026
GLM-5.3$1.791$6.6326$0.17331.0MAug 14, 2026
DeepSeek OCR 2$0.8897$1.4675$0.09248kJan 27, 2026
GPT OSS 120B$0.2311$0.7511$0.0462128kAug 5, 2025
Voxtral Mini 3B$0.0046free—32kJul 1, 2025
Qwen3 Embedding 4B$0.1502free—32kJun 6, 2025
Whisper Large v3$0.0162free—448Oct 1, 2024