Released08/2026
Speed37 tokens/s
Time to first token1.82s
Input price$2/1M
Output price$6/1M
Rankings
Review
Qwen3.8 2.4T A95B excels in reasoning and coding, making it well suited for computationally intensive logic and software development tasks. Deployments must factor in cost and latency, particularly if the primary use case is general chat or autonomous agents.
Strengths
- Exceptional reasoning capabilities, ranking among the top 10 on AIM
- Strong programming skills, well suited for code generation and debugging
- Massive 2.4T scale providing deep contextual comprehension and broad multilingual coverage
Weaknesses
- General chat is not industry-leading, with occasional unnatural phrasing
- AI agent performance is limited and cannot fully replace specialized autonomous systems
- High compute requirements result in larger deployment costs than lighter models
Use cases
Complex data analysis and logical deduction for research applicationsAutomated code generation, optimization, and bug fixing for software projectsDrafting technical documentation and reports requiring deep reasoning and broad context