LLM Economics
#23

Qwen3.6 Plus

Alibaba · Proprietary · Reasoning · 1M context

63/ 100 composite

Released 2026-04-02

Benchmark Scores by Category

Coding
51.9
Knowledge
58.3
Reasoning
85.4
Agentic
53.0
Math
63.0
Multimodal
67.5

Scores based on weighted benchmarks: SWE-bench, LiveCodeBench (Coding); HLE, MMLU-Pro (Knowledge); ARC-AGI-2, LongBench v2 (Reasoning); Terminal-Bench, BrowseComp (Agentic); FrontierMath, AIME (Math).

API Pricing

API pricing not publicly available for this model.

Estimate Your Costs

Use our cost calculators to estimate monthly spend based on your specific workload: