BenchLM comparison
ChatGPT vs Gemini (July 2026): Benchmarks, Price & Verdict
ChatGPT leads the public ranking comparison: GPT-5.6 Sol scores 81.96 at #3, versus Gemini 3.5 Flash at 64.75 and #33. Gemini remains cheaper at the flagship API tier.
This is a comparison of ChatGPT and Gemini as products, so we anchor it to the flagship models behind each app: ChatGPT runs on OpenAI's GPT-5.6 family — Sol, Terra, and Luna(GA on July 9, 2026), while Gemini runs on Google's Gemini 3.1 Pro and Gemini 3.5 Flash. Every number below is BenchLM's sourced benchmark or pricing data for those models.
Search it as "ChatGPT vs Gemini" or "Gemini vs ChatGPT" — either way, this is the highest-volume rivalry in consumer AI, and in July 2026 the data gives a more interesting answer than the market-share numbers suggest.
The short answer
The public ranking puts GPT-5.6 Sol ahead at 81.96, with Gemini 3.5 Flash at 64.75. The two rows sit at #3 and #33, respectively.
ChatGPT changed its engine with the GPT-5.6 family on July 9, 2026. Sol now has a public Estimated score, and its sourced results are strong precisely where Gemini 3.1 Pro is weakest: 91.9 on Terminal-Bench 2 versus Gemini's 77, and 92.2 on BrowseComp versus 86. If agentic work is your use case, the verdict may flip once Sol is fully scored.
On price, Google is the aggressor: Gemini 3.1 Pro costs $2 input / $12 output per million tokens against $5 / $30 for GPT-5.6 Sol. ChatGPT leads the current public score; Gemini's case is lower API cost.
ChatGPT vs Gemini: the models behind the apps
Here are the flagship models side by side. Public scores and ranks come from the main leaderboard lane; Arena Elo is shown separately where published.
| Model | BenchLM score | Arena Elo | Context |
|---|---|---|---|
| GPT-5.6 SolOpenAI | 81.96 | 1486.4 | 1M |
| GPT-5.6 TerraOpenAI | 72.57 | — | 1M |
| GPT-5.6 LunaOpenAI | 67.17 | — | 1M |
| GPT-5.5OpenAI | 73.51 | 1481.0 | 1M |
| Gemini 3.1 ProGoogle | 55.3 | 1485.6 | 1M |
| Gemini 3.5 FlashGoogle | 64.75 | 1476.3 | 1M |
For the benchmark-by-benchmark detail, see the dedicated model pages: Gemini 3.1 Pro vs GPT-5.6 Sol for the current flagship matchup, or Gemini 3.1 Pro vs GPT-5.5 for the fully scored comparison.
ChatGPT vs Gemini pricing: subscriptions and API
ChatGPT Plus costs $20/monthand is the reference consumer AI subscription. Google sells Gemini access through its paid Google AI subscription tiers, which bundle the Gemini app's top models with Workspace and storage perks — a different shape of deal that is hard to compare line-for-line, and Google also gives away more capability at the free tier than OpenAI does.
The API numbers are directly comparable. Google wins decisively at the flagship tier — Gemini 3.1 Pro costs less than half of GPT-5.6 Sol per token — while at the budget tier OpenAI's GPT-5.6 Luna ($1 / $6) actually undercuts Gemini 3.5 Flash ($1.50 / $9). Full rate cards live on the OpenAI API pricing and Google API pricing hubs.
| Model | Input / 1M tokens | Output / 1M tokens | Note |
|---|---|---|---|
| GPT-5.6 Sol | $5.00 | $30.00 | — |
| GPT-5.6 Terra | $2.50 | $15.00 | — |
| GPT-5.6 Luna | $1.00 | $6.00 | — |
| GPT-5.5 | $5.00 | $30.00 | Previous ChatGPT flagship, still on the API |
| Gemini 3.1 Pro | $2.00 | $12.00 | Rises to $4 / $18 above 200K-token prompts |
| Gemini 3.5 Flash | $1.50 | $9.00 | — |
Which should you pick?
Pick ChatGPT if…
- Agentic and terminal work is the job: GPT-5.6 Sol's early 91.9 on Terminal-Bench 2 crushes Gemini 3.1 Pro's 77.
- You lean on web-research agents — Sol's 92.2 BrowseComp leads Gemini 3.1 Pro's 86.
- You want the newest flagship family: GPT-5.6 shipped July 9, 2026, a full generation fresher than Gemini 3.1 Pro.
- The ChatGPT ecosystem matters — custom GPTs, voice, and the most mature plugin and memory surface in consumer AI.
- You want a predictable $20/month subscription rather than a bundle tied to Google's ecosystem.
Pick Gemini if…
- Price matters at the top end: $2 / $12 per million tokens for Gemini 3.1 Pro versus $5 / $30 for GPT-5.6 Sol.
- You want a fast model that still ranks near the frontier: Gemini 3.5 Flash scores 64.75 at #33.
- Human preference is your proxy for quality — Gemini 3.1 Pro's 1485.8 Arena Elo tops every scored OpenAI model on this page.
- You live in Google's ecosystem: Gemini is wired into Gmail, Docs, and Android in ways ChatGPT cannot match.
- You want strong math and reasoning on a budget — 3.1 Pro posts 97 on MATH-500 and 92 on BBH.
How this comparison works
BenchLM compares products through the models that power them — "ChatGPT" and "Gemini" each swap in new flagship models several times a year, so a product verdict is only as current as the models behind it. Overall scores use the current public ranking lane; Arena Elo is the community human-preference rating where one has been published; API prices come from each provider's official rate card as recorded in BenchLM's pricing data. Models outside the public ranking show as unranked instead of receiving a legacy score. Every individual benchmark quoted in the text comes from the same sourced dataset that powers our model comparison pages.
ChatGPT vs Gemini FAQ
Is Gemini better than ChatGPT?
Not on the current public score. GPT-5.6 Sol scores 81.96 at #3, while Gemini 3.5 Flash scores 64.75 at #33. Gemini can still be the better purchase when lower API pricing matters more than the point-score lead.
Is Gemini cheaper than ChatGPT?
At the flagship tier, yes: Gemini 3.1 Pro costs $2 input / $12 output per million tokens — less than half of GPT-5.6 Sol's $5 / $30. At the budget tier the order flips: GPT-5.6 Luna ($1 / $6) is cheaper than Gemini 3.5 Flash ($1.50 / $9). On subscriptions, ChatGPT Plus is a flat $20/month while Google bundles Gemini into its paid AI tiers, with a generous free tier below.
Which is better for coding, ChatGPT or Gemini?
ChatGPT, on the current sourced data. GPT-5.6 Sol's early coding results — 91.9 on Terminal-Bench 2 and a 77.4 Artificial Analysis coding index — lead Gemini 3.1 Pro's 77 and 68.8 on the same measures, and even the older GPT-5.5 beats it on Terminal-Bench 2 (82 vs 77). Gemini 3.1 Pro fights back on SWE-bench Verified (75) and shines on math-heavy work.
Can I use both ChatGPT and Gemini?
Absolutely — and Gemini's free tier makes this the cheapest "both" combo of the big three. A sensible split: keep ChatGPT Plus for agentic tasks, browsing research, and its ecosystem, and use Gemini free or paid for long-document work, math, and anything living in Gmail or Docs. On the API, routing bulk workloads to Gemini 3.5 Flash while reserving Sol for hard agentic jobs is a common cost play.
Related comparisons
- ChatGPT vs ClaudeGPT-5.6 against Claude Fable 5 on benchmarks and price.
- Claude vs GeminiThe closest flagship matchup on the leaderboard.
- All model comparisonsCompare any two models BenchLM tracks, benchmark by benchmark.
- OpenAI API pricingEvery GPT model with current per-token rates.
- Google API pricingEvery Gemini model with current per-token rates.