Slide deck summaries

Claude Fable 5 vs Gemini 3.6 Flash
for slide deck summaries

This page compares two current models on one job: turning ordered slide screenshots into a written summary a stakeholder can read without opening the deck. It looks at numerical accuracy, cost and a fair way to test both on your own decks.

Sep 1, 2026 · 10 min read

The bottom line
Fable 5 for the numbers

Claude Fable 5 is the safer default when keeping the actual numbers straight is the deciding requirement. Gemini 3.6 Flash is the better economic choice for bulk extraction and acceptable first-pass quality.

That split rests on the closest public professional-document benchmark, where Fable 5 more than doubles Gemini 3.6 Flash's strict pass rate1, set against Gemini's published rate, which runs at a fraction of Fable's on both input and output5. Fable's own benchmark result still means most of that deliberately difficult test was not passed perfectly, so even the stronger model needs a validation step.

A practical two-stage workflow uses Gemini 3.6 Flash for inexpensive bulk extraction or triage, then Fable 5 to reconcile conflicting numbers and write the final stakeholder narrative. Even with Fable, do not publish without a number ledger and a source-slide check.

Who this is for
Which reporting roles this fits

Start with Fable 501

Finance and investor relations

A wrong figure in front of investors is expensive. Fable 5's lead on the closest professional-document benchmark supports it as the safer default for numerical fidelity.

Use Gemini at scale02

Sales and account teams

You summarize dozens of decks a week for internal updates. Gemini 3.6 Flash's low cost and chart-reading evidence suit a fast, economical first pass.

Triage then finish03

Strategy and consulting

Some decks matter more than others. Let Gemini triage the bulk of the deck library, then route the decision-critical ones to Fable 5 for the final numbers-checked summary.

Verify before sending04

Operations reporting

A dashboard screenshot feeds a real decision. Require a source-slide reference for every material number from either model before the summary goes out.

What we compared
The models not the slide viewer

This page compares the two models through their API in one neutral setup, not one model inside a presentation tool against the other inside a document viewer.

The parts that matter for a slide-deck summary are reading small labels and footnotes, extracting values from charts and tables, keeping a repeated metric consistent across slides, distinguishing actuals from forecasts, and turning the facts into readable prose. Official docs come first, then independent professional-document and chart-reasoning benchmarks with a clear method.

We left presentation-software and file-conversion features out of the spec table on purpose. Whether a screenshot comes from a slide viewer, an export tool or a browser extension belongs to the app around the model, not to the model itself. Judging those here would compare tools, not which model keeps the numbers straight.

Specs at a glance
The screenshot-relevant numbers

The model facts that actually affect a slide-deck summary. Presentation and file-conversion tools are left out, since they belong to the app around the model.

Spec
Claude Fable 5
Gemini 3.6 Flash
Why it matters
Context window
1,000,000 tokens
1,048,576 tokens
Either can normally hold an entire ordinary deck in one request29
Max output
128,000 tokens
65,536 tokens
Fable 5 can return a much longer written summary in one pass29
Inputs
Text and image, with a higher-resolution tier for small labels and footnotes, up to 600 images per request
Text, image, audio, video and PDF, with per-image media-resolution control
Gemini reads more input types directly. Fable's higher-resolution tier suits dense charts and tiny annotations91011
List price
$10 in / $50 out per million
$0.75 in / $3.75 out per million through Dec 31 2026, rising to $1.50 in / $7.50 out from Jan 1 2027
Gemini is far cheaper for bulk processing, especially at its current promotional rate25
Reasoning effort
Adaptive thinking always enabled, effort adjustable
Thinking levels from minimal through high, medium as default
Higher effort helps reconcile repeated figures across slides and costs more on both sides212

Figures from Anthropic and Google documentation, checked September 1, 2026. Gemini's rate is promotional through December 31, 2026, after which Google's published rate rises.

Head to head
Where each model leads by dimension

The answer changes by dimension, not by brand. This is the main analysis: which model has the edge on each part of turning slide screenshots into a summary, and what backs it up.

Dimension
Better choice
Why the edge exists
Best evidence
Keeping numbers and supporting facts straight
Claude Fable 5
The strongest public evidence for this task, though it tests native PDFs with tables and charts rather than ordered screenshots
29.8% strict pass rate against 14.0% for Gemini 3.6 Flash at high reasoning on GDP.pdf's current leaderboard1
Dense charts and small annotations
Claude Fable 5, provisionally
Anthropic reports improved accuracy on detailed technical figures, and its higher-resolution image tier suits tiny labels. Partly vendor evidence, so treat it as directional
Vendor guidance on precise number extraction from scientific figures, plus a higher-resolution image tier1013
Complex chart reasoning
Too close to call from like-for-like data
Google's own chart-reasoning result is strong, but published Fable figures on similar tests use different harnesses, so neither should decide this row
85.2% without tools and 89.4% with tools for Gemini 3.6 Flash on CharXiv Reasoning3
Following a stakeholder-summary brief
Claude Fable 5, qualitative
Anthropic's own guidance describes stronger instruction following for outcome-first, complete-sentence summaries. No independent like-for-like benchmark was found
Anthropic's model-specific prompting guidance for financial analysis and document summaries4
Very long or numerous decks
Tie in practice
Both accept roughly one million tokens. Gemini's own long-context retrieval figures are strong but have no comparable Fable 5 result in the same table
91.8% at 128K and 54.0% at 1M tokens on Google's own eight-needle long-context test, no Fable 5 figure in the same table3
Cost and bulk processing
Gemini 3.6 Flash
Its published standard rate is a fraction of Fable's on both input and output, and batch pricing lowers it further
$0.75 in / $3.75 out per million against Fable's $10 in / $50 out, through Dec 31 20265
Current model position
Claude Fable 5
Fable remains Anthropic's highest-capability generally available model. Google now calls Gemini 3.6 Flash its previous-generation Flash model, following the August 2026 release of 3.7 Flash
Gemini 3.6 Flash listed as previous-generation after Gemini 3.7 Flash's release6

Better-choice calls map to dimensions the sources actually evaluated. Visual-token figures are not directly comparable across vendors, since they use different image encoders and token accounting.

How to test
A fair test on your own decks

A useful test feels boring. Same ordered screenshots, same prompt, same output limit. Then judge what your team actually pays for: every material number kept straight, and a summary that reads without the deck.

Sample01

Pick three to five real decks

Include a dense KPI dashboard, a financial deck with actual, target and forecast columns, a chart-heavy strategy deck with small labels, a deck where the same metric repeats, and a deliberately low-resolution set.

Prompt02

Give both the same prompt

Use the same ordered screenshots and output limit, and require a reasoning level appropriate to each model. Do not edit the results before scoring.

Setup03

Use the same setup

Build a gold-standard ledger of every material number: value, sign, currency, unit, percentage versus percentage points, period, scenario and source slide.

Scoring04

Score without editing first

Check exact numerical matches, missing or invented numbers, correct attribution to actual, forecast or target, and readability for someone who never saw the deck. Blind the reviewers for commercial work.

What the evidence shows
Directional and not yet settled

No public benchmark exactly measures ordered slide-deck screenshots to a standalone summary. Here is what each source helps judge, and how much weight it can carry.

Source
What it measures
What it suggests
How to weigh it
GDP.pdf leaderboard
100 professional PDF question-document pairs, strict all-criteria pass
Fable 5 scores 29.8% against 14.0% for Gemini 3.6 Flash at high reasoning
The best available proxy, though it uses native PDFs rather than ordered screenshots1
Google's own later model table
The same style of professional-document benchmark, a different harness
Reports 22.0% for Gemini 3.6 Flash, against Surge's current 14.0%
Configuration and harness details materially affect results, so neither figure is a universal error rate6
CharXiv Reasoning (Google)
Synthesizing information from complex scientific charts
Gemini 3.6 Flash reaches 85.2% without tools and 89.4% with tools
Supports Gemini for chart-heavy extraction, but does not test multi-slide numeric consistency3
Independent medical-report study
A narrowly structured feature-extraction task on ultrasound reports
Gemini 3.6 Flash slightly ahead of Fable 5, 98.9% against 97.8%
Useful evidence Gemini can win a constrained extraction workflow, not evidence it wins screenshot chart reading7

GDP.pdf's documented failure patterns include misaligned tables, misread charts, skipped exclusions and amendments overriding earlier text, all close analogues to a slide summary that quotes the right-looking number from the wrong series or period.

How to prompt each one
A ledger before the narrative

The best prompt is not the same for both. Fable 5 benefits from a precise deliverable and an explicit ban on inferring unreadable values. Gemini benefits from splitting extraction from prose with structured output.

For Claude Fable 5, give it a precise deliverable, require a fact ledger before synthesis, and explicitly request brevity. High effort is the documented starting point, moving to medium only after testing numerical accuracy.

For Gemini 3.6 Flash, use high media resolution for dense slides, consider high thinking, and split extraction from prose using structured output so the narrative pass cannot silently change a number the extraction pass already recorded.

A Claude Fable 5 prompt: a ledger before the narrative

Review the slides in order. First extract every material number
with its unit, period, series, and slide number. Reconcile
repeated metrics before writing.

Then produce a 500-word stakeholder summary. Never infer an
unreadable value, write "unreadable on slide N" instead.

Lead with the business outcome and use complete sentences.

A Gemini 3.6 Flash prompt: structured extraction first

For each slide, return JSON fields: slide, claim, value, unit,
period, series, qualifier, confidence. Copy numbers exactly, do
not calculate or round. Use null when unreadable.

After completing all slides, compare repeated metrics and write a
concise stakeholder summary using only the extracted records.

Weak spots
And how to fix them

Neither model is a safe unsupervised summarizer. The useful question is where each one adds risk, and what to change in the prompt or the workflow.

Model
Weak spot
What it looks like
How to fix it
Claude Fable 5
High price, and higher effort may over-deliberate
Produces more detail than a stakeholder needs, and can still misread a poor or tiny image despite the higher-resolution tier.
Add a hard word limit and an outcome-first instruction. Pre-resize screenshots without destroying text. Require an intermediate fact ledger4.
Gemini 3.6 Flash
Lower result on the closest current professional-document benchmark
Low or default media resolution may lose fine print, and it is no longer Google's newest Flash model.
Set dense slides to high resolution and thinking to high. Extract structured facts first, then run a separate consistency pass. Include 3.7 Flash in acceptance testing16.
Both
Can produce plausible prose that hides a wrong source value
A swapped series, a missing minus sign, or a changed unit inside an otherwise fluent summary.
Require source-slide references for every material number. Reject any output containing a number absent from the ledger. Use deterministic validation for totals and period labels.

Which one to choose
Start with the cost of a wrong number

One question first. Is a wrong number materially worse than paying more for inference? Then follow the branch that matches your deck volume and audience.

Is a wrong number worse than paying more? Yes, numerical risk is high No, volume or unit cost dominates Thousands of slides, some critical Mostly narrative, few numbers Deployment not yet built Claude Fable 5 Gemini 3.6 Flash Gemini to triage, Fable 5 to finish Gemini 3.6 Flash Test Gemini 3.7 Flash alongside

A starting point, not a rule. Test on your own decks before you commit.

Recommendations
Pick by numerical risk and volume

If a wrong number is materially worse than paying more for inference, start with Claude Fable 5 for extraction, reconciliation and final writing. Use high-quality lossless images and high effort for dense charts or footnotes, and require human verification of the number ledger before it reaches an executive or investor14.

If cost or volume dominates, start with Gemini 3.6 Flash. Use high media resolution only on complex slides, use structured output for the extraction phase, and upgrade the final pass to high thinking35.

If there are thousands of slides but only some are decision-critical, use Gemini for triage and Fable 5 for the flagged slides and the final summary. If the deployment has not yet been built, test Gemini 3.7 Flash alongside these two, since 3.6 is already previous-generation6.

Bottom line
Fable 5 for the numbers

Claude Fable 5 is the safer default for turning slide screenshots into a stakeholder-ready summary when keeping the actual numbers straight is the deciding requirement. Gemini 3.6 Flash is the better economic choice and has credible chart and long-context capabilities.

No public benchmark exactly measures ordered slide-deck screenshots against a standalone executive summary. GDP.pdf is the best available proxy, but it uses PDFs, and reported Gemini scores differ between evaluation harnesses. Image preprocessing can also alter results, and Gemini's introductory price expires on December 31, 2026. Playgram is not the right buy for everyone either: a solo analyst who only ever needs one model is better served by a single vendor subscription.

The safest final step is to test the shape of your own decks, not a generic prompt from the internet. A fair test needs the same setup for both models: the same screenshots, the same prompt and the same place to run them, so the result reflects the models and not the tool around them. In practice that is harder than it sounds, since most teams end up running one model in one app and the other in a different one, on two separate subscriptions, which tilts the comparison before the first summary comes back. The cleaner the setup, the more the difference you see is really Claude Fable 5 vs Gemini 3.6 Flash, and not just which one happened to be easier to reach that day.

Test both in one workspace
Right here inside Playgram

That's the practical case for the setup just described, and it's also how the day-to-day work gets easier. When both models sit in one workspace, you can send the same slide screenshots to each, compare the summaries side by side, and hand a deck from one model to the other without setting it up again.

Try it on a deck your team is reviewing right now. Paste the ordered screenshots in once, put the same request in front of the latest GPT and Claude models, and keep the conversation going with whichever summary keeps the numbers straighter instead of starting over for a second opinion.

The same memory carries across the team too, not just this one comparison, over the latest GPT, Claude, Gemini and Grok models and many more, all in one place8. The line-up is curated, so retired models are turned off and new ones are added as they ship.

Team memory

Shared across everyone and every model.

Project memory

Scoped to a campaign or document set.

Personal memory

Your own working style, kept private.

Fair pricing
Pay per usage, not per seat

Upgrade as needed, and only pay for what you actually use

Save ~17% with the annual plan

Pro

$50/ month

Perfect for small and medium teams

Unlimited users & infinite memory

Multi-LLM chats

Granular access control to models

EU data residency

Get started

Ultra

$200/ month

Best for large, growing teams

Unlimited users & infinite memory

Multi-LLM chats

Unlimited use of DeepSeek V4 Flash

Granular access control to models

Choose US or EU data residency

Get started

Enterprise

Get in touch

Unlimited Credits

For organizations with advanced needs

Unlimited users & SSO

Priority Support

Unlimited use of DeepSeek V4 Flash

Granular access control to models

Choose US or EU data residency

Book a call

30-days money back guarantee

Frequently asked
questions

Claude Fable 5, on the closest public evidence. On GDP.pdf's current leaderboard, a strict all-criteria benchmark over professional documents with tables, charts and footnotes, Fable 5 scores a 29.8% pass rate against 14.0% for Gemini 3.6 Flash at high reasoning. The benchmark uses native PDFs rather than ordered screenshots, so it is the best available proxy rather than a direct test.

No. Google released Gemini 3.7 Flash in August 2026 and now calls 3.6 its previous-generation Flash model. Gemini 3.6 Flash remains stable with no announced shutdown date, so a decision pinned to it today stays valid, but a new procurement test should include 3.7 Flash as a candidate.

Yes, on its own published evidence. Google reports 85.2% without tools and 89.4% with tools on CharXiv Reasoning, a test of synthesizing information from complex scientific charts. That supports Gemini for chart-heavy extraction, but it does not test whether a multi-slide summary preserves every period, unit and qualification the way GDP.pdf does.

Through December 31, 2026, Gemini 3.6 Flash publishes $0.75 per million input tokens and $3.75 per million output, against Fable 5's $10 and $50. Gemini also offers batch pricing at $0.375 per million input and $1.875 per million output during the introductory period, roughly half its standard rate. Google's own published rate for Gemini rises to $1.50 and $7.50 from January 1, 2027.

No. Run the same prompt on both and compare the answers, or switch between them mid-conversation. You choose after reading both answers instead of guessing up front.

Related comparisons

Claude Sonnet 5 vs Gemini 3.6 FlashClaude Opus 5 vs Gemini 3.6 Flash for finance memosGrok 4.5 vs Gemini 3.6 Flash for competitor battlecardsClaude Opus 5 vs Gemini 3.1 Pro for chart screenshots

One deck for both models
One plan for the whole team

Send the same slide screenshots to the latest GPT and Claude models, keep the number ledger in one place, and see which summary needs less checking. Set it up in a minute.

Get startedSee the pricing