ChatGPT vs Gemini 3 vs Grok Make Fortnite From Scratch
By Minimunch · Tech · 1M views · 13:37
The teardown in brief
What's working
- Concept lands within 8 seconds and the three-AI format is immediately clear — viewers know exactly what they're watching and why it's interesting before the first prompt is sent
- Gemini segment delivers genuine escalating payoffs (movement fix → vehicles → inventory → building) that keep momentum building through the video's strongest 5 minutes
- The Victory Royale payoffs for both ChatGPT and Grok create natural closure for each segment and give the viewer a satisfying 'chapter complete' beat
What's costing attention
- No failure consequence established — the challenge lacks stakes beyond 'one AI is better,' which means the debugging sections feel like neutral filler rather than tense obstacles
- Grok 4.1 narrated recap breaks the video's core promise (watching AI fail in real time) and hands viewers an exit ramp at the worst possible moment — 8 minutes in, just as the final round starts
- Audio delivery is too flat for this audience — the video's three biggest moments (cars working, Gemini building, final ranking) all happen at the same energy level as routine commentary
The first 30 seconds
Chat GPT, Gemini 3, and Grok. Today, we're putting these AIs against each other to see who can make the best Fortnite clone from scratch in 30 minutes. We're going to be starting off with Chat GPT 5.1. I had Chat GPT make me an actual decent prompt, so the results we get should be a lot better. Let's go ahead and send
Hook fires at 6 seconds with the full premise crystal-clear — three named AIs, same challenge, 30 minutes each. Concept reaffirms the title immediately. The one miss is zero stakes established, so the packaging drop is softened by hook clarity but not by emotional investment.
Where viewers drop
8:10 — Grok 4.1 Narrated Recap (critical)
You spend 40 seconds narrating a bunch of failures that happened off-screen — 10 failed iterations, misspelled code, random images — none of which the viewer gets to see. You're describing entertainment instead of showing it.
Why it matters — Viewers clicked to watch AI fail in real time, not to hear a summary. This is the one moment in the video where the format completely breaks — the promise of the content vanishes and you're just talking at them.
0:00 — No Stakes Established — Curiosity Only (critical)
You set up the whole premise with zero consequences for any AI. The viewer knows three AIs will be tested and one will 'win,' but there's nothing at stake if they fail — no consequence, no bet, no penalty, no raised bar. Every debugging session becomes emotionally neutral because nothing bad can actually happen.
Why it matters — Without a failure consequence, every time an AI breaks, it's mildly funny instead of genuinely tense. With stakes, every bug becomes something the viewer fears. Right now you're leaning entirely on curiosity, which bleeds out by the Grok segment.
10:00 — Grok Debugging Treadmill (moderate)
Over roughly 100 seconds, Grok produces four lines of code after thinking for 83 seconds, then the page goes unresponsive twice, and the creator takes a break. Three consecutive 'it didn't work' moments with no visible progress and no payoff in between.
Why it matters — After the Grok 4.1 narrated recap already deflated expectations for this segment, hitting three more dead ends in a row with no laughs, no reveals, and no forward momentum tells viewers the Grok chapter is just going to be frustrating to watch.
3:35 — Low-Energy Delivery Throughout (moderate)
The entire Gemini segment — where the genuinely impressive things happen (cars, inventory, building, winning) — sits at -24 to -28dB, mostly in the quiet-to-normal range. Your audio energy is at the same level whether you're describing a broken camera or discovering working vehicles.
Why it matters — For a gaming challenge audience, flat delivery during the coolest moments flattens the emotional payoffs. When cars appear at 5:40 and you're at -26dB saying 'I'm just so impressed,' viewers feel the gap between what the moment deserves and what they're receiving.
How the video is built
- 0:00 ChatGPT Round — Build, Break, Barely Win
- 3:21 Gemini 3 Round — Dominant Performance
- 8:10 Grok Round — Chaos and Salvage
- 12:41 Final Ranking and Comparison
What any creator can steal
- Add a failure consequence to the hook
- Cut or replace the Grok 4.1 narrated recap (8:09–8:45)
- Add visible timer shots at each AI transition
- Add forward bridges at both AI transitions instead of clean breaks
- Raise your audio energy at the three biggest payoff moments
- Before you start filming, write down one specific failure consequence — what happens to you if the worst AI wins? 'I stream it for an hour' or 'I have to submit it as a game jam entry' costs you nothing to set up and transforms the entire video's emotional register.
More teardowns from Minimunch
Want this on your own video?
Paste any YouTube URL and Retti maps every drop, spike and plateau to the moment that caused it.
Analyse a video free