retti.ai › Teardowns › ChatGPT vs Gemini 3 vs Grok Make Fortnite From Scratch
Predicted Retention Teardown

ChatGPT vs Gemini 3 vs Grok Make Fortnite From Scratch

By Minimunch · Tech · 1M views · 13:37

ChatGPT vs Gemini 3 vs Grok Make Fortnite From Scratch

The teardown in brief

What's working

What's costing attention

The first 30 seconds

Chat GPT, Gemini 3, and Grok. Today, we're putting these AIs against each other to see who can make the best Fortnite clone from scratch in 30 minutes. We're going to be starting off with Chat GPT 5.1. I had Chat GPT make me an actual decent prompt, so the results we get should be a lot better. Let's go ahead and send

Hook fires at 6 seconds with the full premise crystal-clear — three named AIs, same challenge, 30 minutes each. Concept reaffirms the title immediately. The one miss is zero stakes established, so the packaging drop is softened by hook clarity but not by emotional investment.

Where viewers drop

8:10 — Grok 4.1 Narrated Recap (critical)

You spend 40 seconds narrating a bunch of failures that happened off-screen — 10 failed iterations, misspelled code, random images — none of which the viewer gets to see. You're describing entertainment instead of showing it.

Why it matters — Viewers clicked to watch AI fail in real time, not to hear a summary. This is the one moment in the video where the format completely breaks — the promise of the content vanishes and you're just talking at them.

0:00 — No Stakes Established — Curiosity Only (critical)

You set up the whole premise with zero consequences for any AI. The viewer knows three AIs will be tested and one will 'win,' but there's nothing at stake if they fail — no consequence, no bet, no penalty, no raised bar. Every debugging session becomes emotionally neutral because nothing bad can actually happen.

Why it matters — Without a failure consequence, every time an AI breaks, it's mildly funny instead of genuinely tense. With stakes, every bug becomes something the viewer fears. Right now you're leaning entirely on curiosity, which bleeds out by the Grok segment.

10:00 — Grok Debugging Treadmill (moderate)

Over roughly 100 seconds, Grok produces four lines of code after thinking for 83 seconds, then the page goes unresponsive twice, and the creator takes a break. Three consecutive 'it didn't work' moments with no visible progress and no payoff in between.

Why it matters — After the Grok 4.1 narrated recap already deflated expectations for this segment, hitting three more dead ends in a row with no laughs, no reveals, and no forward momentum tells viewers the Grok chapter is just going to be frustrating to watch.

3:35 — Low-Energy Delivery Throughout (moderate)

The entire Gemini segment — where the genuinely impressive things happen (cars, inventory, building, winning) — sits at -24 to -28dB, mostly in the quiet-to-normal range. Your audio energy is at the same level whether you're describing a broken camera or discovering working vehicles.

Why it matters — For a gaming challenge audience, flat delivery during the coolest moments flattens the emotional payoffs. When cars appear at 5:40 and you're at -26dB saying 'I'm just so impressed,' viewers feel the gap between what the moment deserves and what they're receiving.

How the video is built

What any creator can steal

More teardowns from Minimunch

Want this on your own video?

Paste any YouTube URL and Retti maps every drop, spike and plateau to the moment that caused it.

Analyse a video free