Independent & unofficial. Not affiliated with Anthropic. Facts verified 21 August 2026. Always confirm pricing at claude.com/pricing.
Comparison

Claude vs Gemini in August 2026

Claude and Google Gemini compared on the things that differ most - long-context retrieval, speed, multimodal generation and price - with verified August 2026 figures.

Claude and Gemini are not competing on the same axis. As of August 2026, Google's flagship is Gemini 3.7 Flash - released 13 August 2026 - and it beats Claude decisively on long-context retrieval, raw speed and price, while Anthropic's Opus 5 and Fable 5 hold the top of the independent intelligence and human-preference rankings. Google also generates images, video and music. Claude generates none.

The two clearest Google wins come first on this page, because they are the ones most comparison articles bury.

#Head to head, August 2026

Google flagship
Gemini 3.7 Flash - a Flash model, released 13 August 2026
Google Pro line
Gemini 3.1 Pro, still in preview
Anthropic lineup
Fable 5, Opus 5, Sonnet 5, Haiku 4.5
Long-context retrieval
GDM-MRCR v2 - Gemini 97.0% vs Sonnet 5 81.5%
Output speed
Gemini 3.7 Flash ~3,901 tok/s vs Fable 5 ~75 tok/s
API price per MTok
Gemini 3.7 Flash $0.75/$3.75 vs Fable 5 $10/$50

#Where does Gemini beat Claude?

#1. Long-context retrieval - a 1M window is not the same as reliable recall

This is the most important finding on the page, and it cuts against the marketing on both sides. Claude offers a 1,000,000-token context window as the default on Fable 5, Opus 5 and Sonnet 5, with no beta header and no long-context price premium. That is a genuine engineering achievement. But window size measures what you can load, not what the model can reliably find again.

On GDM-MRCR v2, a multi-round needle-retrieval benchmark, Google's own Gemini 3.7 Flash model card reports 97.0% against Claude Sonnet 5's 81.5% and GPT-5.6 Terra's 93.5%. On LVBench, a long-video benchmark, it is 85.4% against Sonnet 5's 68.5% - though that comparison is partly moot, since Claude cannot accept video input at all. Claude has published very little long-context retrieval evidence of its own, and loses the retrieval benchmarks that do exist. The one exception in the published record runs the other way: on GraphWalks BFS at 1M tokens, OpenAI's launch table gives Anthropic's Mythos 5 79.4% against GPT-5.6 Sol's 77.1%.

Practical reading: if your workload is "drop 600 pages in and ask precise questions about page 412", Gemini is the safer tool. If it is "reason carefully across a codebase you have structured yourself", the retrieval gap matters much less. Our model comparison works through which Claude model to use when context length is the binding constraint.

#2. Speed and price - Google wins by an order of magnitude

Artificial Analysis measures Gemini 3.7 Flash at roughly 3,901 output tokens per second. Fable 5 runs at roughly 75. That is not a tuning difference; it is a different class of product. Opus 5 is much better at around 626 tokens/sec, but still six times slower.

52×faster: Gemini 3.7 Flash vs Fable 5 output speed
13×cheaper: $0.75/$3.75 vs $10/$50 per MTok
7Intelligence Index points behind Opus 5 (56 vs 63)

Gemini 3.7 Flash's $0.75 in / $3.75 out is introductory pricing through 31 December 2026, and Google's Gemini API pricing page states rates double to $1.50/$7.50 on 1 January 2027. Even after that doubling it undercuts Sonnet 5 at $2/$10. Anthropic's counter-levers are the Batch API's flat 50% discount and prompt caching at 0.1× base input on cache reads - both covered on the Claude API page - but they apply to Google's rates too.

#3. Hard-reasoning breadth

On Humanity's Last Exam (Verified), Google's card reports Gemini 3.7 Flash at 53.6% against Claude Sonnet 5's 31.0% - a 22-point gap against a model costing roughly a quarter as much. Anthropic has not published an Opus 5 or Fable 5 figure on that benchmark, so the flagship-to-flagship comparison does not exist. Treat the 22 points as a mid-tier result, not a settled verdict on the whole family, but note that the absence of an Anthropic number is itself part of the problem.

#4. Generative media, which Claude simply does not do

Google ships Veo 3.1 for video, Imagen for images, Lyria 3 for music, Flow and Flow Music for production workflows, and Gemini Omni for multimodal creation. Gemini 3.7 Flash accepts text, image, audio and video as input. Claude accepts text and images and emits text only - no image, video, audio or music generation, and no video or audio input. For any creative or media pipeline this is decisive, and no benchmark changes it. The same gap against OpenAI is covered on our ChatGPT comparison.

#5. Distribution

Gemini is the default assistant in Google Search's AI Mode, in Gmail, in Docs and on Android, plus NotebookLM and a bundled consumer stack. Anthropic has no search engine, no mobile OS and no office suite. Claude reaches you because you went looking for it; Gemini reaches you because you opened your inbox.

#What is each company actually shipping?

One structural fact is worth pausing on: Google's best model is a Flash model. The Pro line has not shipped a stable release past Gemini 3.1 Pro, which remains in preview. Google describes 3.7 Flash as its "most intelligent workhorse model yet for coding and agents" - that is a company competing on intelligence-per-dollar, not on an absolute frontier model. Anthropic is doing the opposite.

ModelStatusContext / outputInput / output per MTokInputs accepted
Gemini 3.7 FlashStable, flagship1M / 64k$0.75 / $3.75 (introductory)Text, image, audio, video
Gemini 3.1 ProPreviewNot documented on a current card$2.00–$4.00 / $12.00–$18.00 by contextMultimodal
Claude Fable 5GA 9 Jun 20261M / 128k$10 / $50Text, image
Claude Opus 5GA 24 Jul 20261M / 128k$5 / $25Text, image
Claude Sonnet 5GA 30 Jun 20261M / 128k$2 / $10Text, image
Claude Haiku 4.5GA 1 Oct 2025200k / 64k$1 / $5Text, image

Gemini 3.7 Flash's knowledge cutoff is March 2026 (January 2025 in some domains). Among Claude models, Opus 5 has the newest at May 2026 - cutoffs are not in version order, since Fable 5 and Sonnet 5 both stop at January 2026. The full specification set is on the Claude models index.

#Where does Claude beat Gemini?

Four areas, and they are not consolation prizes.

MeasureClaudeGeminiSource
Artificial Analysis Intelligence Index63 (Opus 5, max) - #156 (3.7 Flash, high) - #16Artificial Analysis, independent
arena.ai Elo (human preference)1507 (Fable 5) - #1, six of top twelve1490 (3.7 Flash High) - #9; 3.1 Pro Preview 1483arena.ai, independent
ARC-AGI-330.16% (Opus 5, high effort)No comparable published resultARC Prize, independent
GDPval-AA v2 (Elo)1598 (Sonnet 5)1525 (3.7 Flash)Google's own model card
Agents' Last Exam33.3% (Sonnet 5)26.3% (3.7 Flash)Google's own model card
Terminal-Bench 2.183.8% (Claude Code + Fable 5) - top of the official board65.8% (Gemini CLI + 3.1 Pro)Official Terminal-Bench leaderboard
BioMysteryBench (human-solvable)87.5% (Sonnet 5)87.1% (3.7 Flash)Google's own model card

Two things make that table trustworthy. First, several of those Claude wins appear on Google's own model card - Google did not cherry-pick a sweep, which is to its credit. Second, the Intelligence Index and Elo figures come from third parties running their own harness rather than from either vendor.

But read Google's card carefully in the other direction too: it compares Gemini 3.7 Flash against Sonnet 5 and GPT-5.6 Terra, not against Fable 5 or Opus 5. Gemini beats Sonnet 5 on twelve of eighteen rows at roughly a quarter of the price, which is a real result - but it is a mid-tier comparison being read as a flagship one. Anthropic makes the mirror-image mistake: its Fable 5 benchmark table is published only as an image, and the Opus 5 post gives relative claims rather than absolute numbers, so the flagship-to-flagship table nobody can build is partly Anthropic's fault.

#Agentic coding is still Claude's strongest ground

Claude Code plus Fable 5 leads the official Terminal-Bench 2.1 board at 83.8%; no Gemini model appears near the top of Terminal-Bench or SWE-bench. Where Gemini does compete on code, it competes on value - it edges Sonnet 5 on FrontierCode 1.1 (43.6% vs 42.7%) and leads Code Arena Web Elo (1588 vs 1541) at a quarter of the cost. That is a strong argument for high-volume code work and a weak one for the hardest single task in your backlog. The tooling side is covered on Claude for developers.

#Enterprise compliance depth

Anthropic publishes Enterprise at a flat $20/seat/month plus usage at API rates, with SCIM, audit logs, a Compliance API, custom data retention, customer-managed encryption keys, US-only inference, IP allowlisting and HIPAA-ready configuration listed openly. Google's strength is different in kind: it wins where the customer already lives in Workspace and GCP, with Gemini Enterprise Agent Platform as the deployment path. If your procurement process is a control checklist rather than an existing contract, Anthropic's published surface is easier to evaluate.

#How do the consumer plans compare?

One price we cannot pin down

Google AI Plus is reported at both $4.99 and $7.99 per month by credible third-party trackers, and Google's own plan pages do not render prices to automated fetchers. One source also disagrees with Google on the storage allowance. Treat the entry tier as roughly $5–$8 and confirm it yourself at one.google.com before budgeting.

TierClaudeGoogleNote
Free$0, metered weekly pool, 5-hour rolling refresh$0, Gemini 3.6 Flash with a 32k context ceilingGoogle's free context cap is the tightest of the big three
Entry paidNothing offeredAI Plus, ~$5–$8 (see caveat)Anthropic has nothing between $0 and $20
Flagship individualPro, $20/mo or $200/yr (~$17/mo)AI Pro, $19.99/moGoogle bundles Veo 3.1, Lyria 3, Deep Research and 5 TB storage
Heavy individualMax 5× $100 / Max 20× $200AI Ultra $99.99 (5×) / $199.99 (20×)Effectively at parity
BusinessTeam from $20/seat/mo annualBundled into Workspace tiersDifferent shapes; compare on what you already own

Claude plan prices are from claude.com/pricing and explained in detail on the pricing page; what the free tier actually includes is on the Claude free plan page. Note that Google's $19.99 tier ships generative video and music inside the subscription, which Claude Pro cannot match at any price.

#Claude or Gemini: which should you pick?

If your job is…PickBecause
Question-answering over very large document setsGeminiGDM-MRCR v2 97.0% vs 81.5%; retrieval, not window size, is the constraint
Anything involving video, audio or musicGeminiVeo 3.1, Lyria 3, and video/audio input; Claude has none of it
High-volume, latency-sensitive production trafficGemini~3,901 tok/s at $0.75/$3.75
Long-running agentic coding on a real repositoryClaudeTop of the official Terminal-Bench 2.1 board; no Gemini model near it
Novel reasoning with no template to imitateClaudeARC-AGI-3 at 30.16%, independently verified by ARC Prize
Drafting and editing where tone mattersClaudeSix of the top twelve arena.ai slots, Fable 5 at #1
Regulated buying against a published control listClaudeFlat published Enterprise pricing with the compliance surface itemised
You already run Workspace and GCPGeminiDistribution and procurement beat a seven-point index gap

If neither answer fits, the wider field is surveyed on Claude alternatives - including open-weight models now within a few Intelligence Index points of the frontier. Our overall verdict on whether Claude is worth its price is on the Claude review page, and the OpenAI comparison is on Claude vs ChatGPT.

#Frequently asked questions

Does Claude or Gemini handle long documents better?

Gemini, on the evidence available. Both offer a one-million-token context window, but on the GDM-MRCR v2 retrieval benchmark Google reports Gemini 3.7 Flash at 97.0% against Claude Sonnet 5's 81.5%. A large window measures what you can load, not what the model can reliably find again.

Is Gemini cheaper than Claude?

Substantially. Gemini 3.7 Flash costs $0.75 in and $3.75 out per million tokens against Fable 5's $10 and $50 - roughly thirteen times cheaper. Those Gemini rates are introductory through 31 December 2026 and double on 1 January 2027, and still undercut Claude Sonnet 5.

Why is Google's flagship a Flash model?

Because Google is competing on intelligence-per-dollar rather than absolute capability. Gemini 3.7 Flash shipped on 13 August 2026 as the stable flagship while the Pro line has not moved past Gemini 3.1 Pro, still in preview. Google positions 3.7 Flash as a workhorse for coding and agents.

Can Claude do anything Gemini cannot?

Yes. Claude Opus 5 leads the independent Artificial Analysis Intelligence Index at 63 against Gemini 3.7 Flash's 56, Fable 5 leads arena.ai human preference at 1507, and Claude Code with Fable 5 tops the official Terminal-Bench 2.1 coding board where no Gemini model ranks near the top.

Can Claude generate video or music like Gemini?

No. Claude generates text and code only, and cannot accept video or audio as input. Google ships Veo 3.1 for video, Imagen for images and Lyria 3 for music, all reachable from paid Gemini subscriptions. For media work Claude is not a candidate.

How much does Google AI Plus cost?

Reported figures disagree. Credible trackers give both $4.99 and $7.99 per month, and Google's own plan pages do not render prices to automated tools. Treat the entry tier as roughly $5 to $8 and confirm directly with Google before budgeting. Anthropic offers no equivalent tier.

Verify it yourself

Claude figures verified 21 August 2026 against claude.com/pricing and platform.claude.com/docs. Gemini benchmark and specification figures come from the Gemini 3.7 Flash model card and API rates from ai.google.dev. Independent rankings are from Artificial Analysis, arena.ai and ARC Prize. Both vendors ship frequently - re-check before committing.