Skip to content
Sovyron

Pareto Inference

1 modelsProvider docs

As of Oct 4, 2026, Pareto Inference lists 1 supported model in the Sovyron catalog, including GLM-5.3-Flash — glm-flash families. Its metered rates sit at $0.0475 per 1M tokens on a 3:1 input:output blend, and every price in the table is Pareto Inference's own, normalized to USD — open a model to compare it with the cheapest offer across all providers.

Supported models and pricing

1 of 1 models
Pareto Inference model prices in USD per 1M tokens, updated Oct 4, 2026
Model1M input1M output1M cache readContextReleased
GLM-5.3-Flash$0.03$0.1$0.0061.0MAug 26, 2026