Claude vs Gemini in August 2026
Claude and Google Gemini compared on the things that differ most - long-context retrieval, speed, multimodal generation and price - with verified August 2026 figures.
Claude and Gemini are not competing on the same axis. As of August 2026, Google's flagship is Gemini 3.7 Flash - released 13 August 2026 - and it beats Claude decisively on long-context retrieval, raw speed and price, while Anthropic's Opus 5 and Fable 5 hold the top of the independent intelligence and human-preference rankings. Google also generates images, video and music. Claude generates none.
The two clearest Google wins come first on this page, because they are the ones most comparison articles bury.
#Head to head, August 2026
- Google flagship
- Gemini 3.7 Flash - a Flash model, released 13 August 2026
- Google Pro line
- Gemini 3.1 Pro, still in preview
- Anthropic lineup
- Fable 5, Opus 5, Sonnet 5, Haiku 4.5
- Long-context retrieval
- GDM-MRCR v2 - Gemini 97.0% vs Sonnet 5 81.5%
- Output speed
- Gemini 3.7 Flash ~3,901 tok/s vs Fable 5 ~75 tok/s
- API price per MTok
- Gemini 3.7 Flash $0.75/$3.75 vs Fable 5 $10/$50
#Where does Gemini beat Claude?
#1. Long-context retrieval - a 1M window is not the same as reliable recall
This is the most important finding on the page, and it cuts against the marketing on both sides. Claude offers a 1,000,000-token context window as the default on Fable 5, Opus 5 and Sonnet 5, with no beta header and no long-context price premium. That is a genuine engineering achievement. But window size measures what you can load, not what the model can reliably find again.
On GDM-MRCR v2, a multi-round needle-retrieval benchmark, Google's own Gemini 3.7 Flash model card reports 97.0% against Claude Sonnet 5's 81.5% and GPT-5.6 Terra's 93.5%. On LVBench, a long-video benchmark, it is 85.4% against Sonnet 5's 68.5% - though that comparison is partly moot, since Claude cannot accept video input at all. Claude has published very little long-context retrieval evidence of its own, and loses the retrieval benchmarks that do exist. The one exception in the published record runs the other way: on GraphWalks BFS at 1M tokens, OpenAI's launch table gives Anthropic's Mythos 5 79.4% against GPT-5.6 Sol's 77.1%.
Practical reading: if your workload is "drop 600 pages in and ask precise questions about page 412", Gemini is the safer tool. If it is "reason carefully across a codebase you have structured yourself", the retrieval gap matters much less. Our model comparison works through which Claude model to use when context length is the binding constraint.
#2. Speed and price - Google wins by an order of magnitude
Artificial Analysis measures Gemini 3.7 Flash at roughly 3,901 output tokens per second. Fable 5 runs at roughly 75. That is not a tuning difference; it is a different class of product. Opus 5 is much better at around 626 tokens/sec, but still six times slower.
Gemini 3.7 Flash's $0.75 in / $3.75 out is introductory pricing through 31 December 2026, and Google's Gemini API pricing page states rates double to $1.50/$7.50 on 1 January 2027. Even after that doubling it undercuts Sonnet 5 at $2/$10. Anthropic's counter-levers are the Batch API's flat 50% discount and prompt caching at 0.1× base input on cache reads - both covered on the Claude API page - but they apply to Google's rates too.
#3. Hard-reasoning breadth
On Humanity's Last Exam (Verified), Google's card reports Gemini 3.7 Flash at 53.6% against Claude Sonnet 5's 31.0% - a 22-point gap against a model costing roughly a quarter as much. Anthropic has not published an Opus 5 or Fable 5 figure on that benchmark, so the flagship-to-flagship comparison does not exist. Treat the 22 points as a mid-tier result, not a settled verdict on the whole family, but note that the absence of an Anthropic number is itself part of the problem.
#4. Generative media, which Claude simply does not do
Google ships Veo 3.1 for video, Imagen for images, Lyria 3 for music, Flow and Flow Music for production workflows, and Gemini Omni for multimodal creation. Gemini 3.7 Flash accepts text, image, audio and video as input. Claude accepts text and images and emits text only - no image, video, audio or music generation, and no video or audio input. For any creative or media pipeline this is decisive, and no benchmark changes it. The same gap against OpenAI is covered on our ChatGPT comparison.
#5. Distribution
Gemini is the default assistant in Google Search's AI Mode, in Gmail, in Docs and on Android, plus NotebookLM and a bundled consumer stack. Anthropic has no search engine, no mobile OS and no office suite. Claude reaches you because you went looking for it; Gemini reaches you because you opened your inbox.
#What is each company actually shipping?
One structural fact is worth pausing on: Google's best model is a Flash model. The Pro line has not shipped a stable release past Gemini 3.1 Pro, which remains in preview. Google describes 3.7 Flash as its "most intelligent workhorse model yet for coding and agents" - that is a company competing on intelligence-per-dollar, not on an absolute frontier model. Anthropic is doing the opposite.
| Model | Status | Context / output | Input / output per MTok | Inputs accepted |
|---|---|---|---|---|
| Gemini 3.7 Flash | Stable, flagship | 1M / 64k | $0.75 / $3.75 (introductory) | Text, image, audio, video |
| Gemini 3.1 Pro | Preview | Not documented on a current card | $2.00–$4.00 / $12.00–$18.00 by context | Multimodal |
| Claude Fable 5 | GA 9 Jun 2026 | 1M / 128k | $10 / $50 | Text, image |
| Claude Opus 5 | GA 24 Jul 2026 | 1M / 128k | $5 / $25 | Text, image |
| Claude Sonnet 5 | GA 30 Jun 2026 | 1M / 128k | $2 / $10 | Text, image |
| Claude Haiku 4.5 | GA 1 Oct 2025 | 200k / 64k | $1 / $5 | Text, image |
Gemini 3.7 Flash's knowledge cutoff is March 2026 (January 2025 in some domains). Among Claude models, Opus 5 has the newest at May 2026 - cutoffs are not in version order, since Fable 5 and Sonnet 5 both stop at January 2026. The full specification set is on the Claude models index.
#Where does Claude beat Gemini?
Four areas, and they are not consolation prizes.
| Measure | Claude | Gemini | Source |
|---|---|---|---|
| Artificial Analysis Intelligence Index | 63 (Opus 5, max) - #1 | 56 (3.7 Flash, high) - #16 | Artificial Analysis, independent |
| arena.ai Elo (human preference) | 1507 (Fable 5) - #1, six of top twelve | 1490 (3.7 Flash High) - #9; 3.1 Pro Preview 1483 | arena.ai, independent |
| ARC-AGI-3 | 30.16% (Opus 5, high effort) | No comparable published result | ARC Prize, independent |
| GDPval-AA v2 (Elo) | 1598 (Sonnet 5) | 1525 (3.7 Flash) | Google's own model card |
| Agents' Last Exam | 33.3% (Sonnet 5) | 26.3% (3.7 Flash) | Google's own model card |
| Terminal-Bench 2.1 | 83.8% (Claude Code + Fable 5) - top of the official board | 65.8% (Gemini CLI + 3.1 Pro) | Official Terminal-Bench leaderboard |
| BioMysteryBench (human-solvable) | 87.5% (Sonnet 5) | 87.1% (3.7 Flash) | Google's own model card |
Two things make that table trustworthy. First, several of those Claude wins appear on Google's own model card - Google did not cherry-pick a sweep, which is to its credit. Second, the Intelligence Index and Elo figures come from third parties running their own harness rather than from either vendor.
But read Google's card carefully in the other direction too: it compares Gemini 3.7 Flash against Sonnet 5 and GPT-5.6 Terra, not against Fable 5 or Opus 5. Gemini beats Sonnet 5 on twelve of eighteen rows at roughly a quarter of the price, which is a real result - but it is a mid-tier comparison being read as a flagship one. Anthropic makes the mirror-image mistake: its Fable 5 benchmark table is published only as an image, and the Opus 5 post gives relative claims rather than absolute numbers, so the flagship-to-flagship table nobody can build is partly Anthropic's fault.
#Agentic coding is still Claude's strongest ground
Claude Code plus Fable 5 leads the official Terminal-Bench 2.1 board at 83.8%; no Gemini model appears near the top of Terminal-Bench or SWE-bench. Where Gemini does compete on code, it competes on value - it edges Sonnet 5 on FrontierCode 1.1 (43.6% vs 42.7%) and leads Code Arena Web Elo (1588 vs 1541) at a quarter of the cost. That is a strong argument for high-volume code work and a weak one for the hardest single task in your backlog. The tooling side is covered on Claude for developers.
#Enterprise compliance depth
Anthropic publishes Enterprise at a flat $20/seat/month plus usage at API rates, with SCIM, audit logs, a Compliance API, custom data retention, customer-managed encryption keys, US-only inference, IP allowlisting and HIPAA-ready configuration listed openly. Google's strength is different in kind: it wins where the customer already lives in Workspace and GCP, with Gemini Enterprise Agent Platform as the deployment path. If your procurement process is a control checklist rather than an existing contract, Anthropic's published surface is easier to evaluate.
#How do the consumer plans compare?
Google AI Plus is reported at both $4.99 and $7.99 per month by credible third-party trackers, and Google's own plan pages do not render prices to automated fetchers. One source also disagrees with Google on the storage allowance. Treat the entry tier as roughly $5–$8 and confirm it yourself at one.google.com before budgeting.
| Tier | Claude | Note | |
|---|---|---|---|
| Free | $0, metered weekly pool, 5-hour rolling refresh | $0, Gemini 3.6 Flash with a 32k context ceiling | Google's free context cap is the tightest of the big three |
| Entry paid | Nothing offered | AI Plus, ~$5–$8 (see caveat) | Anthropic has nothing between $0 and $20 |
| Flagship individual | Pro, $20/mo or $200/yr (~$17/mo) | AI Pro, $19.99/mo | Google bundles Veo 3.1, Lyria 3, Deep Research and 5 TB storage |
| Heavy individual | Max 5× $100 / Max 20× $200 | AI Ultra $99.99 (5×) / $199.99 (20×) | Effectively at parity |
| Business | Team from $20/seat/mo annual | Bundled into Workspace tiers | Different shapes; compare on what you already own |
Claude plan prices are from claude.com/pricing and explained in detail on the pricing page; what the free tier actually includes is on the Claude free plan page. Note that Google's $19.99 tier ships generative video and music inside the subscription, which Claude Pro cannot match at any price.
#Claude or Gemini: which should you pick?
| If your job is… | Pick | Because |
|---|---|---|
| Question-answering over very large document sets | Gemini | GDM-MRCR v2 97.0% vs 81.5%; retrieval, not window size, is the constraint |
| Anything involving video, audio or music | Gemini | Veo 3.1, Lyria 3, and video/audio input; Claude has none of it |
| High-volume, latency-sensitive production traffic | Gemini | ~3,901 tok/s at $0.75/$3.75 |
| Long-running agentic coding on a real repository | Claude | Top of the official Terminal-Bench 2.1 board; no Gemini model near it |
| Novel reasoning with no template to imitate | Claude | ARC-AGI-3 at 30.16%, independently verified by ARC Prize |
| Drafting and editing where tone matters | Claude | Six of the top twelve arena.ai slots, Fable 5 at #1 |
| Regulated buying against a published control list | Claude | Flat published Enterprise pricing with the compliance surface itemised |
| You already run Workspace and GCP | Gemini | Distribution and procurement beat a seven-point index gap |
If neither answer fits, the wider field is surveyed on Claude alternatives - including open-weight models now within a few Intelligence Index points of the frontier. Our overall verdict on whether Claude is worth its price is on the Claude review page, and the OpenAI comparison is on Claude vs ChatGPT.
#Frequently asked questions
Does Claude or Gemini handle long documents better?
Gemini, on the evidence available. Both offer a one-million-token context window, but on the GDM-MRCR v2 retrieval benchmark Google reports Gemini 3.7 Flash at 97.0% against Claude Sonnet 5's 81.5%. A large window measures what you can load, not what the model can reliably find again.
Is Gemini cheaper than Claude?
Substantially. Gemini 3.7 Flash costs $0.75 in and $3.75 out per million tokens against Fable 5's $10 and $50 - roughly thirteen times cheaper. Those Gemini rates are introductory through 31 December 2026 and double on 1 January 2027, and still undercut Claude Sonnet 5.
Why is Google's flagship a Flash model?
Because Google is competing on intelligence-per-dollar rather than absolute capability. Gemini 3.7 Flash shipped on 13 August 2026 as the stable flagship while the Pro line has not moved past Gemini 3.1 Pro, still in preview. Google positions 3.7 Flash as a workhorse for coding and agents.
Can Claude do anything Gemini cannot?
Yes. Claude Opus 5 leads the independent Artificial Analysis Intelligence Index at 63 against Gemini 3.7 Flash's 56, Fable 5 leads arena.ai human preference at 1507, and Claude Code with Fable 5 tops the official Terminal-Bench 2.1 coding board where no Gemini model ranks near the top.
Can Claude generate video or music like Gemini?
No. Claude generates text and code only, and cannot accept video or audio as input. Google ships Veo 3.1 for video, Imagen for images and Lyria 3 for music, all reachable from paid Gemini subscriptions. For media work Claude is not a candidate.
How much does Google AI Plus cost?
Reported figures disagree. Credible trackers give both $4.99 and $7.99 per month, and Google's own plan pages do not render prices to automated tools. Treat the entry tier as roughly $5 to $8 and confirm directly with Google before budgeting. Anthropic offers no equivalent tier.
Claude figures verified 21 August 2026 against claude.com/pricing and platform.claude.com/docs. Gemini benchmark and specification figures come from the Gemini 3.7 Flash model card and API rates from ai.google.dev. Independent rankings are from Artificial Analysis, arena.ai and ARC Prize. Both vendors ship frequently - re-check before committing.