Model comparison
Claude Mythos 5 vs Exaone 4.0 32B
Head-to-head evidence from 0 shared benchmark results across 0 categories. Overall scores shown here use the public BenchAlign v5 ranking lane.
Public leaderboard positions: Claude Mythos 5 #1 (Supported); Exaone 4.0 32B #170 (Estimated). Intervals and evidence labels describe ranking uncertainty, not a guarantee for a specific workload.
Evidence parity. Claude Mythos 5 and Exaone 4.0 32B share 0 comparable benchmark results. 1 of 8 categories are comparable. 15 results are unique to Claude Mythos 5; 13 to Exaone 4.0 32B.
Updated July 23, 2026- Shared results
- 0
- Claude Mythos 5 only
- 15
- Exaone 4.0 32B only
- 13
- Comparable categories
- 1 / 8
Pick Claude Mythos 5 if you want the stronger benchmark profile. Exaone 4.0 32B only becomes the better choice if knowledge is the priority.
Confidence note. This is a partial-evidence comparison with 0 shared benchmark results across 0 evidence categories; 1 of 8 categories currently have scoreable aggregates for both models. Treat the verdict as directional until coverage is more balanced.
Why this result
Claude Mythos 5 is clearly ahead on the BenchAlign aggregate, 83.93 to 40.44. The gap is large enough that you do not need to squint at the spreadsheet to see the difference.
Claude Mythos 5 gives you the larger context window at 1M+, compared with 128K for Exaone 4.0 32B.
Operational comparison
Runtime and commercial metrics are compared only when both models have a complete sourced value.
| Metric | Claude Mythos 5 | Exaone 4.0 32B | Comparison |
|---|---|---|---|
| Input / output priceUSD per 1M tokens | Claude Mythos 5$10 input / $50 output | Exaone 4.0 32BNot available | A complete price comparison is not available. |
| Generation speedtokens per second | Claude Mythos 5Not available | Exaone 4.0 32BNot available | A complete speed comparison is not available. |
| First-answer latencyseconds to first token | Claude Mythos 5Not available | Exaone 4.0 32BNot available | A complete latency comparison is not available. |
| Context windowmaximum listed tokens | Claude Mythos 51M+ | Exaone 4.0 32B128K | Claude Mythos 5 lists the larger context window. |
Benchmark Deep Dive
Agentic5 benchmarks
Coding4 benchmarks
Reasoning2 benchmarks
KnowledgeExaone 4.0 32B wins10 benchmarks
| Benchmark | Claude Mythos 5 | Exaone 4.0 32B | Result |
|---|---|---|---|
| GPQASource | 94.1% | — | Not comparable |
| HLESource | 64.5% | — | Not comparable |
| HLE w/o toolsSource | 59% | — | Not comparable |
| MMLU-ProSource | — | 81.8% | Not comparable |
| Artificial Analysis Intelligence IndexSource | — | 6.0% | Not comparable |
| AA-GPQA DiamondSource | — | 62.8% | Not comparable |
| AA-HLESource | — | 4.9% | Not comparable |
| AA-Omniscience IndexSource | — | -62.3% | Not comparable |
| AA-Omniscience AccuracySource | — | 10.4% | Not comparable |
| AA-Omniscience Hallucination RateSource | — | 81.0% | Not comparable |
Math2 benchmarks
Multilingual1 benchmarks
| Benchmark | Claude Mythos 5 | Exaone 4.0 32B | Result |
|---|---|---|---|
| SWE MultilingualSource | 92.2% | — | Not comparable |
Multimodal3 benchmarks
Inst. Following1 benchmarks
| Benchmark | Claude Mythos 5 | Exaone 4.0 32B | Result |
|---|---|---|---|
| AA-IFBenchSource | — | 33.5% | Not comparable |
Frequently Asked Questions (2)
Which is better, Claude Mythos 5 or Exaone 4.0 32B?
Claude Mythos 5 is ahead on BenchLM's BenchAlign leaderboard, 83.93 to 40.44.
Which is better for knowledge tasks, Claude Mythos 5 or Exaone 4.0 32B?
Exaone 4.0 32B has the edge for knowledge tasks in this comparison, averaging 81.8 versus 68.5. Claude Mythos 5 stays close enough that the answer can still flip depending on your workload.
Related Comparisons
Choose a model with this week’s evidence
Join 2,000+ readers for ranking moves, pricing changes, and the claims that still need proof.
One email each week. Unsubscribe anytime.