Skip to main content

BenchLM recommendation

Best Non-Reasoning LLMs in 2026

Data verified

As of July 23, 2026, the top model in best non-reasoning llms on the BenchLM leaderboard is Claude Opus 4.7 with a score of 71.9.

Last verified: July 23, 2026

Top standard AI models (no chain-of-thought reasoning) ranked by benchmark performance. Faster and cheaper than reasoning models.

Unless noted otherwise, ranking surfaces on this page use BenchLM's provisional leaderboard lane rather than the stricter sourced-only verified leaderboard.

Bottom line: Non-reasoning models are faster and cheaper than chain-of-thought alternatives. Gemini 3.1 Pro leads this tier — proving that strong reasoning scores are possible without dedicated thinking tokens.

Claude Opus 4.7 leads this ranking with a score of 71.9, followed by MiniMax M3 (69.8) and Claude Opus 4.6 (68.6). There is meaningful separation between the top models, suggesting genuine performance differences.

The best open-weight option is MiniMax M3 (ranked #2 with a score of 69.8). Open-weight models are highly competitive in this category — self-hosting is a viable alternative to proprietary APIs.

This ranking is based on provisional overall weighted scores across BenchLM.ai's scoring formula tracked by BenchLM.ai. For detailed model profiles, click any model name below. To compare two specific models head-to-head, use the "vs #" links.

What changed

Gemini 3.1 Pro leads non-reasoning models — best reasoning (97), knowledge (96), and multilingual (100).

Claude Opus 4.6 most consistent non-reasoning model across all 8 categories.

Claude Sonnet 4.6 strong mid-tier with best multimodal (95) in this tier.

How to choose

Full Rankings (92 models)

1
Claude Opus 4.7
Anthropic·Proprietary·1M

71.9

BenchAlign v5

2
MiniMax M3
MiniMax·Open Weight·1M

69.8

BenchAlign v5

3
Claude Opus 4.6
Anthropic·Proprietary·1M

68.6

BenchAlign v5

4
Gemini 3 Pro
Google·Proprietary·2M

67.7

BenchAlign v5

5
Inkling
Thinking Machines Lab·Open Weight·1M

67.5

BenchAlign v5

6
GLM-5
Z.AI·Open Weight·200K

66.1

BenchAlign v5

7
Claude Sonnet 4.6
Anthropic·Proprietary·200K

65.1

BenchAlign v5

8
Claude Opus 4.5
Anthropic·Proprietary·200K

64.2

BenchAlign v5

9
MiniMax M2.7
MiniMax·Open Weight·200K

64.1

BenchAlign v5

10
GLM-5V-Turbo
Z.AI·Proprietary·200K

63.5

BenchAlign v5

11
DeepSeek V4 Pro
DeepSeek·Open Weight·1M

60.7

BenchAlign v5

12
Gemini 3 Flash
Google·Proprietary·1M

60.5

BenchAlign v5

13
Grok 4
xAI·Proprietary·128K

60.4

BenchAlign v5

14
Grok 4.1
xAI·Proprietary·1M

60

BenchAlign v5

15
Kimi K2.5
Moonshot AI·Open Weight·256K

59.7

BenchAlign v5

16
MiniMax M2.5
MiniMax·Proprietary·128K

59.5

BenchAlign v5

17
DeepSeek V4 Flash
DeepSeek·Open Weight·1M

58.9

BenchAlign v5

18
GLM-4.5
Z.AI·Proprietary·128K

57.6

BenchAlign v5

19
Gemini 2.5 Pro
Google·Proprietary·1M

57.3

BenchAlign v5

20
Qwen3.5 397B
Alibaba·Open Weight·128K

57

BenchAlign v5

21
Claude Haiku 4.5
Anthropic·Proprietary·200K

56.6

BenchAlign v5

22
Qwen3 235B 2507
Alibaba·Open Weight·128K

56

BenchAlign v5

23
Trinity-Large-Preview
Arcee AI·Open Weight·512K

55.9

BenchAlign v5

24
DeepSeek V3.2
DeepSeek·Open Weight·128K

55.4

BenchAlign v5

25
Gemini 3.1 Pro
Google·Proprietary·1M

55.3

BenchAlign v5

26
Step 3.5 Flash
StepFun·Open Weight·256K

55.1

BenchAlign v5

27
DeepSeek LLM 2.0
DeepSeek·Open Weight·128K

54.5

BenchAlign v5

28
DeepSeek V3.1
DeepSeek·Open Weight·128K

53.6

BenchAlign v5

29
Claude Sonnet 4.5
Anthropic·Proprietary·200K

53.6

BenchAlign v5

30
Nemotron 3 Nano 30B
NVIDIA·Open Weight·32K

52.9

BenchAlign v5

31
Qwen2.5-72B
Alibaba·Open Weight·128K

52.2

BenchAlign v5

32
Llama 3.1 405B
Meta·Open Weight·128K

51.7

BenchAlign v5

33
Grok 4.1 Fast
xAI·Proprietary·1M

51.3

BenchAlign v5

34
GPT-4.1
OpenAI·Proprietary·1M

51.1

BenchAlign v5

35
Nemotron 3 Super 120B A12B
NVIDIA·Open Weight·256K

51

BenchAlign v5

36
Gemini 3.1 Flash-Lite
Google·Proprietary·1M

50.8

BenchAlign v5

37
Llama 3 70B
Meta·Open Weight·128K

50.8

BenchAlign v5

38
Mistral Large 3
Mistral·Proprietary·128K

50.4

BenchAlign v5

39
DeepSeek Coder 2.0
DeepSeek·Open Weight·128K

50.3

BenchAlign v5

40
GPT-OSS 120B
OpenAI·Open Weight·128K

50.1

BenchAlign v5

41
Nemotron 3 Super 100B
NVIDIA·Open Weight·1M

50.1

BenchAlign v5

42
Qwen2.5-1M
Alibaba·Open Weight·1M

49.9

BenchAlign v5

43
Seed-2.0-Lite
ByteDance·Proprietary·256K

49.8

BenchAlign v5

44
Aion-2.0
Aion Labs·Proprietary·128K

48.7

BenchAlign v5

45
Mixtral 8x22B Instruct v0.1
Mistral·Open Weight·64K

48.5

BenchAlign v5

46
Gemini 2.5 Flash
Google·Proprietary·1M

48.1

BenchAlign v5

47
Claude 3.5 Sonnet
Anthropic·Proprietary·200K

47.7

BenchAlign v5

48
GLM-4.5-Air
Z.AI·Proprietary·128K

47.7

BenchAlign v5

49
Mistral Small 4
Mistral·Open Weight·256K

46.2

BenchAlign v5

50
Claude 4.1 Opus
Anthropic·Proprietary·200K

45.9

BenchAlign v5

51
Z-1
Z·Proprietary·128K

45.1

BenchAlign v5

52
Nemotron-4 15B
NVIDIA·Open Weight·32K

45.1

BenchAlign v5

53
Mistral 8x7B
Mistral·Open Weight·32K

45

BenchAlign v5

54
DeepSeek V3
DeepSeek·Open Weight·128K

45

BenchAlign v5

55
Moonshot v1
Moonshot AI·Proprietary·128K

44.8

BenchAlign v5

56
Seed-2.0-Mini
ByteDance·Proprietary·256K

44.6

BenchAlign v5

57
GPT-4.1 mini
OpenAI·Proprietary·1M

44.2

BenchAlign v5

58
Ling 2.6 Flash
InclusionAI·Open Weight·262K

43.9

BenchAlign v5

59
Mistral Medium 3
Mistral·Proprietary·128K

43.2

BenchAlign v5

60
Claude 4 Sonnet
Anthropic·Proprietary·200K

42.8

BenchAlign v5

61
GPT-OSS 20B
OpenAI·Open Weight·128K

42.7

BenchAlign v5

62
GPT-4.1 nano
OpenAI·Proprietary·1M

42.1

BenchAlign v5

63
Mistral Large 2
Mistral·Proprietary·128K

41.8

BenchAlign v5

64
Gemma 3 27B
Google·Open Weight·32K

41.6

BenchAlign v5

65
GPT-4o
OpenAI·Proprietary·128K

41.5

BenchAlign v5

66
Claude 3 Opus
Anthropic·Proprietary·200K

41.1

BenchAlign v5

67
Grok 3 [Beta]
xAI·Proprietary·128K

40.4

BenchAlign v5

68
Mistral 7B v0.3
Mistral·Open Weight·32K

39.9

BenchAlign v5

69
Llama 4 Scout
Meta·Open Weight·10M

39.9

BenchAlign v5

70
Qwen2.5-VL-32B
Alibaba·Open Weight·32K

39.9

BenchAlign v5

71
Llama 4 Behemoth
Meta·Open Weight·32K

39.8

BenchAlign v5

72
Mistral 8x7B v0.2
Mistral·Open Weight·32K

39.1

BenchAlign v5

73
Exaone 4.0 1.2B
LG AI Research·Open Weight·128K

39.1

BenchAlign v5

74
Grok Code Fast 1
xAI·Proprietary·256K

38.6

BenchAlign v5

75
Granite-4.0-350M
IBM·Open Weight·32K

38.3

BenchAlign v5

76
Granite-4.0-H-350M
IBM·Open Weight·32K

38.3

BenchAlign v5

77
GPT-4o mini
OpenAI·Proprietary·128K

37.9

BenchAlign v5

78
Gemini 1.5 Pro
Google·Proprietary·2M

35.7

BenchAlign v5

79
Qwen2.5 Coder 32B Instruct
Alibaba·Open Weight·128K

34.7

BenchAlign v5

80
Ministral 3 14B
Mistral·Open Weight·128K

34.5

BenchAlign v5

81
GPT-4 Turbo
OpenAI·Proprietary·128K

27.4

BenchAlign v5

82
Kimi K2
Moonshot AI·Proprietary·128K

27.2

BenchAlign v5

83
MiniMax M1 80k
MiniMax·Proprietary·80K

25.1

BenchAlign v5

84
Llama 4 Maverick
Meta·Open Weight·1M

23.5

BenchAlign v5

85
Phi-4
Microsoft·Open Weight·16K

22.7

BenchAlign v5

86
Gemini 1.0 Pro
Google·Proprietary·32K

21.8

BenchAlign v5

87
Claude 3 Haiku
Anthropic·Proprietary·200K

21.4

BenchAlign v5

88
Ministral 3 8B
Mistral·Open Weight·128K

21

BenchAlign v5

89
Nova Pro
Amazon·Proprietary·128K

20.3

BenchAlign v5

90
LFM2-24B-A2B
LiquidAI·Proprietary·32K

18.9

BenchAlign v5

91
Ministral 3 3B
Mistral·Open Weight·128K

18.3

BenchAlign v5

92
LFM2.5-1.2B-Instruct
LiquidAI·Proprietary·32K

15.5

BenchAlign v5

Key Takeaways

The top model is Claude Opus 4.7 by Anthropic with a BenchAlign v5 score of 71.9 and Supported evidence.

The best open-weight model is MiniMax M3 at position #2.

92 models are included in this ranking.

Score in Context

What these scores mean

Non-reasoning models are standard completion/chat models without dedicated chain-of-thought. They are ranked by the same overall BenchLM score and are typically faster and cheaper per token.

Known limitations

The "non-reasoning" label excludes models with explicit chain-of-thought (like o3, DeepSeek R1). Some non-reasoning models still reason internally — the distinction is about architecture and pricing, not capability.

Last updated: July 23, 2026

Choose a model with this week’s evidence

Join 2,000+ readers for ranking moves, pricing changes, and the claims that still need proof.

One email each week. Unsubscribe anytime.