LLM Economics
#1

Claude Mythos 5

Anthropic · Proprietary · Reasoning · 1M+ context

87/ 100 composite

Released 2026-06-09

Benchmark Scores by Category

Coding
83.1
Knowledge
93.8
Agentic
97.7
Multimodal
86.0

Scores based on weighted benchmarks: SWE-bench, LiveCodeBench (Coding); HLE, MMLU-Pro (Knowledge); ARC-AGI-2, LongBench v2 (Reasoning); Terminal-Bench, BrowseComp (Agentic); FrontierMath, AIME (Math).

API Pricing

Input

$10

per 1M tokens

Output

$50

per 1M tokens

Value Score

1.7

score per $1 output

Estimate Your Costs

Use our cost calculators to estimate monthly spend based on your specific workload: