Capability badges appear only where the provider catalog explicitly lists support. A blank cell means the source is silent, not that the feature is absent.
Listed rates
Price across providers
Input price per million tokens as published by each provider. Bars are drawn from listed rates only — no traffic weighting, since the catalog observes no requests.
Lowest input$2/M
Across 7 priced providers
Median input$2/M
Midpoint of listed rates
Highest input$2.5/M
1.2× the lowest listed rate
Output range$6 – $6.25
Per million output tokens
Deep Infra$2/M
OpenRouter$2/M
Eden AI$2/M
Vercel AI Gateway$2/M
Kilo Gateway$2/M
Hugging Face$2.5/M
Cortecs$2.5/M
Cost calculator
Estimate a workload
$0.00
Excludes taxes, non-token charges, and tiered discounts.
5 providers list the identical $2 input rate, so price alone will not separate them — compare context limits, max output, and capabilities above.
Context limits also differ by provider, from 262.144K to 1.04858M tokens. Compare the provider table above before choosing on price alone.
Specification
Capabilities
Recorded from the source catalog and provider listings.
Every result stays attached to its source, version, metric and harness. Scores from different versions are never merged, and the profile below plots each benchmark against its own population rather than on a shared scale.
ModelBench does not infer quality from price, context size, or model name. When a source publishes a comparable result, it appears here with its version and link.
Answered directly from the stored record — nothing here is generated beyond the catalog's own fields.
What is Qwen3.8 2.4T A95B?
Open-weight sparse MoE (2.4T total, 95B active), the open-weight twin of Qwen3.8 Max for coding, research, complex reasoning, and agentic workflows. It is published by Alibaba Qwen and catalogued here from Models.dev.
How much does Qwen3.8 2.4T A95B cost?
Listed input pricing starts at $2 per million tokens from Deep Infra, rising to $2.5 across 7 listed providers.
What is the context length of Qwen3.8 2.4T A95B?
Qwen3.8 2.4T A95B accepts up to 262.144K tokens of context and returns up to 131.072K output tokens.
Does Qwen3.8 2.4T A95B support tool calling and structured output?
Provider catalogs list support for tool calling, structured output, and reasoning.
Which providers serve Qwen3.8 2.4T A95B?
7 providers list this model: Deep Infra, OpenRouter, Eden AI, Vercel AI Gateway, Kilo Gateway, Hugging Face and 1 more.
Are the weights for Qwen3.8 2.4T A95B open?
Yes. The weights are published and downloadable from Hugging Face under the qwen3.8-max license.
When was Qwen3.8 2.4T A95B released?
The catalog records a release date of 2026-08-12, last verified Aug 17, 2026.