AI Slides

Claude Fable 5.1 vs GPT-6 Astra: Which AI makes better consulting slides?

Updated: Oct 2, 2026
Claude Fable 5.1 vs GPT-6 Astra: Which AI makes better consulting slides?

Fable 5.1 vs GPT-6 Astra: Which is better for PowerPoint slides?

Short answer: Claude Fable 5.1. It made the more consulting-like deck in both rounds and was two to three times faster, but neither model produced a client-ready deck.

  • From "Turn this report into a deck", both models returned 16 slides with descriptive titles and a neutral tone. Fable 5.1 met 4 of our 9 consulting criteria; GPT-6 Astra met 2.
  • When we asked for McKinsey-level detail, Fable 5.1 added proper action titles, a consulting tone and a clear structure (6 of 9). GPT-6 Astra improved its titles but stayed text-heavy (2 of 9).
  • Fable 5.1's best deck still contained made-up content and lacked the chart call-outs and implication-led titles of real consulting slides.
  • Fable 5.1 took about 6 and 12 minutes per round; GPT-6 Astra took about 17 and 24.

The world of AI slides is changing fast and within days of each other in September 2026 the two largest AI labs, Anthropic and OpenAI, released new flagship models that are – among other things – reportedly much better at creating slides. Anthropic launched Claude Fable 5.1 on September 1, and OpenAI began rolling out GPT-6 Astra shortly after. OpenAI describes Astra as its best model for producing slides that are well laid out and "convey key points with a structured narrative".

In our previous article, we tested Fable 5.1 on three typical consulting tasks. This time, we decided to pitch the models against each other to see which outputted the best consulting-level slides. We compared the time each model took, what it cost and, most importantly, how close the output came to a McKinsey or BCG deck. We then asked both models to make their decks more consulting-like and scored the results again.

 

How we tested Fable 5.1 and GPT-6 Astra

We used a real, publicly available industry report: Oil & Gas in Egypt: Market Report, a 52-page report written by Osama Kamal for Norwegian Energy Partners (NORWEP) in August 2019. The report covers Egypt's economy, investment rules, drilling activity, the main operators, production sharing contracts, the Zohr and West Nile Delta gas projects, and a long newsfeed of deals. 

It's a typical input for a market study: rich in data, but with no storyline and no recommendation.

Cover of the Oil & Gas in Egypt Market Report, a 52-page report by NORWEP from August 2019 used as the test input

We ran the same two prompts in each model, both on the High effort setting:

Round 1: Turn this report into a deck

Round 2: Can you make it more consulting-like so it looks like decks that McKinsey or BCG consultants make in terms of level of detail and complexity of slides?

Claude Fable 5.1 GPT-6 Astra
App Claude desktop app ChatGPT desktop app
Effort setting High High
Questions before building Deck length and main audience Which style or template to use
Round 1 output 16-slide .pptx deck 16-slide .pptx deck
Round 2 output 23-slide .pptx deck 16-slide .pptx deck

We scored every deck against the same nine criteria we use for all our AI slide tests: layout, focus, action titles, tone, visual balance, storyline, coherence, client-readiness and overall MBB look and feel:

  1. Does the layout of the slide effectively communicate the key message? A slide's layout should emphasize the data on the slide. It makes a big difference if you use a table vs. a chart vs. a diagram to get a key message across and make it easily understandable. The art of choosing layouts is trained within consulting houses through years of seeing examples and getting feedback.
     
  2. Does each slide manage to only keep information on it that’s relevant for the main message of the slide? Just like it's important to choose a layout that communicates the key message, it's also important to clean the slide for any details or info that make it more difficult to grasp the key message. It's fine adding detailed info (in fact, it's encouraged with techniques like the pyramid principle) but having data on the slide that is not directly connected to the main message makes it confusing.
     
  3. Do the slides have good action titles? Action titles summarize the main message of the slide and are a crucial part of any client deliverables.
     
  4. Is the level and tone of text on the slide consulting-like? Consulting slides are often quite detailed and with a tone that carries a certain point of view or conclusion on whatever the slide says.
     
  5. Is the entire deck visually balanced? This is one of the most important aspects of consulting decks. The balance between layouts needs to be interesting and support the arc of the storyline. So a deck with three tables in a row will fail, while a deck that has a chart, a diagram, a table etc. will not.
     
  6. Does the deck follow a structure that mirrors how a consultant would structure it? Storylining is a key part of what makes consulting decks good. Different types of decks should follow different storylines, just like the audience and setting will impact how the deck is structured.
     
  7. Are the messages and numbers coherent throughout? This is a given.
     
  8. Is the output client-ready? Would we send this deck to a client?
     
  9. Do the slides look and feel like McKinsey or BCG consulting slides overall? Not including overall design choices like colors and fonts, do the slides look and feel like real consulting slides?

We rated each criterion Yes, Partially or No.

 

Round 1: Can Fable 5.1 and GPT-6 Astra turn a PDF report into a PowerPoint deck?

For our first round, we wanted to test a simple prompt to get a baseline of the output. For both models, our prompt was simply "Turn this report into a deck."


Claude Fable 5.1: An okay market brief, but not McKinsey-level

Our first test was with Claude Fable 5.1. We uploaded the PDF report and asked Fable to turn it into a deck.

Before building, Fable 5.1 asked two questions: how detailed the deck should be (it recommended around 15 slides) and who the main audience was. It recommended framing the deck for suppliers and service companies, "as the NORWEP report intends". We accepted both recommendations.

Fable 5.1 asks how detailed the deck should be and recommends an executive summary deck of about 15 slides. It also asks who the main audience is and recommends suppliers and service companies, framed around market entry.

Fable 5.1 returned a 16-slide market briefing for suppliers with an executive summary, a body of slides detailing various analyses from the PDF report, and ending with a suggested focus for suppliers based on its own interpretation.

Fable 5.1 outputted an okay deck with a simple prompt but it is nowhere close to a McKinsey deck.

The output of round 1 from Fable 5.1

Judging the deck against our criteria:

Criterion Rating What we saw
Layout Yes It was fairly good at choosing appropriate layouts like an org-style diagram on who buyers are, a table with up/down arrows on various indicators, and charts that matched the types of data being shown.
Focus Yes
Action titles No Not really. Some slides are proper action titles stating the key message of the slide, but many like “How licensing and production sharing work” are simply descriptive.
Tone No There is both too little text (surprising for LLMs) and too neutral text to be proper consulting slides.
Visual balance Yes It has a good mix of charts, diagrams, and text elements.
Storyline Partially Somewhat. It adds the executive summary and recommendations slides, but the storyline itself is a bit confusing and all over the place. It could have benefitted from structuring the various market analyses in sections similar to the original report.
Coherence Yes
Client-readiness No This looks like something an AI or student created.
MBB look and feel No Not at all. They are too light-weight, unstructured, and neutral to be real McKinsey slides.

All in all, a solid first draft but nowhere near the level of slides produced by houses like McKinsey and BCG.

 

GPT-6 Astra: A basic, text-heavy deck

Egypt oil and gas report PDF uploaded with the prompt Turn this report into a deck, using GPT-6 Astra on High effort

Before building, GPT-6 Astra asked which style the presentation should use. It offered to use an uploaded template or one of its own designs, such as a market trends report or a business review. We chose the blue-and-white Business Review.

GPT-6 Astra asks which presentation style to use: upload a template, Market Trends Report or Business Review

GPT-6 Astra also produced 16 slides that follow the structure of the report: economy, rigs, drilling, production, infrastructure, megaprojects, gas targets, licensing, the players, bid rounds, services and downstream projects.

Astra outputted a very basic deck from our first round prompt.

The output of Astra from the basic round 1 prompt.

The output looks pretty basic, and rating the deck against our consulting criteria it falls short:

Criterion Rating What we saw
Layout No Many of the layouts were text-based tables or boxes and didn’t portray the message and data of the slide in the best way possible.
Focus Yes
Action titles No Its titles are better than completely descriptive titles (e.g., “production-sharing model”) but the titles are still too basic and generic to be proper action titles.
Tone No It reads as a basic summary and doesn’t seem like a consulting slide in either text depth, tone, or conclusions.
Visual balance No It relies too heavily on text.
Storyline Partially Just like Fable, it added an executive summary and recommendations, but the main body still felt ad hoc and jumbled.
Coherence Yes
Client-readiness No This looks like a basic text summary and is too confusing and without clear messages to be client-ready.
MBB look and feel No The slides look like an old-school text summary.

The raw outputs of the two models were fine as a basic summary, but far from an actual consulting deck. Let’s now see if we can move it more towards that.

 

Round 2: Can Fable 5.1 and GPT-6 Astra make McKinsey-style slides?

In the same chats as the previous decks were made, we now asked both models the same follow-up question to push them towards consulting standards:

Follow-up prompt asking Fable 5.1 and GPT-6 Astra to make the deck more consulting-like, like McKinsey or BCG decks

Claude Fable 5.1: Closest to a real MBB deliverable

The second prompt changed Fable 5.1's output completely. It rebuilt the deck as a 23-slide detailed deck with a contents page, numbered sections, section trackers, more in-depth recommendations and implications for the “client”, and much more detailed charts and layouts.

All 23 slides of the Fable 5.1 round 2 deck, with an executive summary, action titles, charts and a phased entry plan

The deck from Fable 5.1 when asked to make it more consulting-like.

Like before, we measured the deck against our consulting criteria:

Criterion Rating What we saw
Layout Yes The layouts are good at communicating the main points.
Focus Yes
Action titles Yes The titles are proper action titles.
Tone Yes
Visual balance Yes The deck is quite well balanced between different slide types.
Storyline Yes It’s a neutral but fine structure.
Coherence Partially There is made up AI nonsense in there that needs to be looked at in detail.
Client-readiness No Not yet. It still looks like an AI-generated deck and has old-school takeaway boxes, a “so what” in the executive summary, and other things that make it feel incomplete.
MBB look and feel No They are the closest of all the four runs, but still not at the level of detail that real consulting slides have. It’s small things like call-outs and annotations on charts, understanding how to draw the implications into the title etc., but the little details make the difference.

This deck is the closest to a real consulting deck and could be lifted with few, but effective tools.

GPT-6 Astra: Slightly better, but nowhere close to a real consulting deck

In our final test, we asked GPT-6 Astra to make the deck more consulting-level. It kept the 16-slide structure it had before, but rebuilt most of the slides.

All 16 slides of the GPT-6 Astra round 2 deck, mostly text-heavy tables with sharper action titles than round 1

The output from Astra when asked for a deck that was closer to consulting-level decks.

Unfortunately, although the visual style became tighter and closer to a consulting look, it still was nowhere close to the McKinsey benchmark:

Criterion Rating What we saw
Layout No The majority of layouts are still text-heavy, table-style slides that could be communicated much better with a mix of charts and diagrams.
Focus No Many of the slides have added information that distracts from the main message both in content and because they are formatted to be eye-catching.
Action titles Yes They are much better than round 1.
Tone Partially Somewhat, although it’s missing the more distinct interpretive tone that real MBB decks have.
Visual balance No It is still largely text.
Storyline Partially Somewhat. The structure is basically unchanged from the first round.
Coherence Yes Pretty much.
Client-readiness No
MBB look and feel No Nowhere close. They look AI-generated and lack the visual balance, depth, and perspective that actual consulting slides have.

The second-round deck from GPT-Astra was very disappointing. It seemed to focus on formatting and action titles, but didn’t change the meat of the slides in terms of layout, density etc.

 

Scorecard: Rating Fable 5.1 vs GPT-6 Astra on 9 consulting criteria

The table below summarises how all four decks scored against our nine criteria. A Yes means the deck met the standard of a McKinsey or BCG deck on that criterion while a Partially means it got some of the way there.

Criterion Fable 5.1: Round 1 Fable 5.1: Round 2 GPT-6 Astra: Round 1 GPT-6 Astra: Round 2
Layout Yes Yes No No
Focus Yes Yes Yes No
Action titles No Yes No Yes
Tone No Yes No Partially
Visual balance Yes Yes No No
Storyline Partially Yes Partially Partially
Coherence Yes Partially Yes Yes
Client-readiness No No No No
MBB look and feel No No No No
Total (Yes / Partially / No) 4 / 1 / 4 6 / 1 / 2 2 / 1 / 6 2 / 2 / 5

Based on out test, three findings stand out;

  • Fable 5.1 is much better at the visuals than GPT-6 Astra: 
    Fable 5.1 chose good and effective layouts from the first prompt e.g., an org-style diagram of who the buyers are, an indicator table with up and down arrows, and charts that match the data. GPT-6 Astra relied on text-heavy tables and boxes in both rounds, so its layout and visual balance scored No each time.
     
  • The second prompt helped Fable 5.1 far more than GPT-6 Astra:
    Fable 5.1 moved from generic deck to something much closer to a real MBB deliverable based on the simple prompt. It rebuilt the entire deck.
    GPT-6 Astra seemed to add some details and adjust text, but did not materially change the deck.
     
  • Fable 5.1's best deck still needs a consultant: 
    It's the closest of the four decks to a real consulting deliverable, but it contains made-up content that has to be checked line by line, old-school takeaway boxes and a "so what" label in the executive summary. It also misses the details that make real MBB slides work: call-outs and annotations on charts, and titles that draw out the implication rather than stating the fact.
     

Time and tokens: How long Fable 5.1 and GPT-6 Astra take to build a deck

We also compared how much it cost and how long it took for all four runs:

Claude Fable 5.1 GPT-6 Astra
Time, round 1 ~6 min (incl. ~20 s answering questions) 17 min 32 s
Time, round 2 ~12 min 24 min 17 s
Tokens written (output), round 1 / round 2 ~42k / ~84k ~25k / ~32k
Total tokens processed, round 1 / round 2 ~3.2M / ~5.3M ~6.7M / ~4.8M
Total tokens processed, both rounds ~8.5M ~11.5M
List API price (input / output per million tokens) $10 / $50 $10 / $50

Fable 5.1 wrote more, roughly 126,000 output tokens across both rounds against about 57,000 for GPT-6 Astra, which fits its larger, more detailed second deck. GPT-6 Astra processed more in total, about 11.5 million tokens against 8.5 million, but almost all of it (about 11.1 million) was cached input.

The two apps count usage slightly differently, so treat the figures as indicative rather than an exact cost comparison. In practice, both models built a full deck in minutes, and the time that matters is the review afterwards. 

Conclusion

Neither Claude Fable 5.1 nor GPT-6 Astra can build consulting-grade slides from a simple prompt. Both returned fine decks, but none that were client-ready or close to real McKinsey or BCG decks.

That being said, Fable 5.1 is the better choice for consulting slides. Its first deck was already more visual than GPT-6 Astra's, and when we asked for consulting-level detail it delivered proper action titles, a consulting tone, and a clear structure. GPT-6 Astra improved its titles but stayed text-heavy and kept the same structure.

However, even Fable 5.1's best deck is a strong first draft, not a finished deliverable. To improve the output, ask explicitly for consulting-level detail, give the model good reference slides, and build on your own slide master. See our tips for getting better decks out of Fable 5.1 in our related article.

Frequently asked questions (FAQs)

Q: Which is better for consulting slides, Fable 5.1 or GPT-6 Astra?

A:

In our test, Claude Fable 5.1 was the far superior slide creator. It was especially good on visual balance and effective layouts.

Q: Is GPT-6 Astra better than Claude Fable 5.1?

A:

No, not for McKinsey or BCG-level consulting slides. In our test, Fable 5.1 produced the stronger deck in both rounds. Caveat, we only tested slide creation, not coding, writing or benchmark performance.

Q: Can ChatGPT (GPT-6 Astra) make a PowerPoint presentation?

A:

Yes. GPT-6 Astra in the ChatGPT desktop app turned a 52-page PDF report into a 16-slide PowerPoint deck, after asking which style or template to use.

The output was a workable PowerPoint file, although we found it lacking against a consulting benchmark.

Q: Can Claude Fable 5.1 make PowerPoint slides?

A:

Yes. Fable 5.1 in the Claude desktop app turned the same 52-page report into a fully working PowerPoint file.

The output was closer to a real consulting deliverable but still lack finishing touches.

Q: Are AI-generated decks client-ready?

A:

No. None of the four decks in our test was client-ready. Even the best deck, from Fable 5.1, contained made-up content and lacked the chart annotations and sharp titles of real consulting slides.

Q: Does asking for a McKinsey-style deck improve the output?

A:

It helps, but not equally. Fable 5.1 adjusted output to be significantly better, while GPT-6 Astra only marginally improved the slides.