#10
Qwen3.7 Plus
Alibaba · Proprietary · Reasoning · 1M context
74/ 100 composite
Released 2026-06-03
Benchmark Scores by Category
Coding
69.3
Knowledge
63.6
Reasoning
89.9
Agentic
66.4
Math
79.7
Multimodal
72.5
Scores based on weighted benchmarks: SWE-bench, LiveCodeBench (Coding); HLE, MMLU-Pro (Knowledge); ARC-AGI-2, LongBench v2 (Reasoning); Terminal-Bench, BrowseComp (Agentic); FrontierMath, AIME (Math).
API Pricing
API pricing not publicly available for this model.
Estimate Your Costs
Use our cost calculators to estimate monthly spend based on your specific workload: