As of Oct 4, 2026, Pareto Inference lists 1 supported model in the Sovyron catalog, including GLM-5.3-Flash — glm-flash families. Its metered rates sit at $0.0475 per 1M tokens on a 3:1 input:output blend, and every price in the table is Pareto Inference's own, normalized to USD — open a model to compare it with the cheapest offer across all providers.
Supported models and pricing
1 of 1 models
Pareto Inference model prices in USD per 1M tokens, updated Oct 4, 2026