リリース2026/04
生成速度136 トークン/秒
初動レイテンシ (TTFT)1.05s
入力料金$0.375/1M
出力料金$2.25/1M
ランキング
レビュー
Qwen3.6 35B A3B is a solid choice for projects requiring multimodal processing and reasoning, but exhibits clear limitations in coding and natural conversation. It is best suited for AI and data teams seeking an economical model for multimodal workloads, rather than a coding assistant or conversational chatbot.
長所
- Robust multimodal performance ranking in the global top 10; handles images, text, and cross-modal interactions effectively
- Solid logical reasoning, suitable for complex analytical tasks and multi-step deduction
- Reliable AI agent performance, well-suited for low-to-medium complexity automation workflows
短所
- Weak coding capabilities; unsuitable for complex code generation or debugging
- Subpar general conversational skills; responses can be disjointed or unnatural in standard dialogue
- Context window and throughput lag behind leading models, limiting its effectiveness for long-context or real-time tasks
ユースケース
Multimodal document and image analysis (OCR, visual recognition, and data extraction from scans)Logical reasoning, complex deduction, and multi-step problem-solvingBasic workflow automation and AI agents for document routing and classification
ガイド・動画
Qwen3.6-35B-A3B is an MoE (Mixture-of-Experts) model from Alibaba featuring 35B total parameters with only 3B active per token. Access it directly via DeepInfra or Roboflow, or download it from Hugging Face for local deployment. It handles programming, code analysis, and multimodal processing (text, images, video). Tip: Use Q4_K_M quantization to run on a 24GB GPU, or deploy quickly via Ollama.