Skip to main content

Benchmark profile

VIBE V2

A display-only MiniMax provider benchmark for end-to-end coding-agent and product-building tasks.

Data verified

Benchmark score on VIBE V2 — July 23, 2026

BenchLM mirrors the published score view for VIBE V2. MiniMax M3 leads the public snapshot at 50.1%. BenchLM does not use these results to rank models overall.

1 modelCodingCurrentDisplay onlyUpdated July 23, 2026

Benchmark score table (1 model)

Score
1
MiniMax M3MiniMax · Open weight
50.1%

About VIBE V2

Year

2026

Tasks

End-to-end coding-agent tasks

Format

Task success rate

Difficulty

Frontier coding-agent workflows

MiniMax reports VIBE V2 in the M3 comparison chart. BenchLM tracks it as a display-only provider row because it is not part of the weighted coding schema.

BenchLM freshness & provenance

Version

VIBE V2 2026

Refresh cadence

Quarterly

Staleness state

Current

Question availability

Public benchmark set

CurrentDisplay only

BenchLM uses freshness metadata to decide whether a benchmark should still be treated as a strong differentiator, a benchmark to watch, or a display-only reference. For the full scoring policy, see the BenchLM methodology page.

FAQ

What does VIBE V2 measure?

A display-only MiniMax provider benchmark for end-to-end coding-agent and product-building tasks.

Which model scores highest on VIBE V2?

MiniMax M3 by MiniMax currently leads with a score of 50.1% on VIBE V2.

How many models are evaluated on VIBE V2?

1 AI models have been evaluated on VIBE V2 on BenchLM.

Last updated: July 23, 2026 · BenchLM version VIBE V2 2026

Choose a model with this week’s evidence

Join 2,000+ readers for ranking moves, pricing changes, and the claims that still need proof.

One email each week. Unsubscribe anytime.