Skip to content
ISSUE #70 • Oct 2, 2026 • 8 MIN READ

Can Sonnet 5.5 Really Build a Full Game From One Prompt? I Tested It.

A few days ago, u/Timisageek posted something on r/ClaudeCode that stopped me mid-scroll: Sonnet 5.5 one-shotting a full kart racing game. Five sub-agents turned a single prompt into a playable game in 67 minutes, with no follow-up messages.

I’ve seen plenty of “one prompt” demos that look great in a clip and fall apart the moment you click something. I’ve built features with sub-agents before, but a complete game from a single message?

That I had to see for myself.

So I typed my own version of the prompt into Claude Code on the web, set Sonnet 5.5 to Max effort, and went to bed. Zero messages while the session ran — no corrections, no course adjustments.

One prompt, then silence.

This is what was waiting in the morning.

Kart Blitz launch trailer showing karts racing through low-poly tracks built by Sonnet 5.5

The game is called Kart Blitz — an arcade-style kart racer where every model, track, and sound effect is generated from code. It took about five hours, and I never touched it once.

Want to try it first? Play Kart Blitz in your browser.

.

.

.

The Prompt

Here’s what I sent:

I need you to launch five sonnet 5.5 sub-agents and help me build a triple A quality game that is a clone of [a classic kart racer]. What I want you to do is I want you to launch these sub-agents, build the game without asking me any questions at all, and use 3JS to build the game.

You can use "playwright-cli" to verify your own work.

You may use "firecrawl" if you need to do any more web research

The game should be frontend only for now. All the persistent data will be stored in local storage.

Please create tasks for this first, and then assign sub agent to each task.

Publish the game to Claude Artifact.

And once you're done, report back to me.
Claude Code on the web showing the Sonnet Kart project with the Full Access cloudplay environment badge, Sonnet 5.5 at Max effort, and the full prompt before the session begins

The core ask came straight from the Reddit post (I’ve swapped out the name of the game it asked to clone).

I added a browser so Sonnet could check its own work visually, made Firecrawl available for web research, told it to plan tasks before launching agents, kept everything frontend-only, and added a step to publish the finished game as a Claude Artifact.

You can find the original prompt in the Reddit thread.

The browser was the most important addition — it would let Sonnet see whether its game actually worked.

Worth knowing: my prompt adds instructions and I ran Max effort where the original used High, so this is a variation on the experiment.

One-shot, as I’m using the word here, means one prompt from me, then silence until the report came back.

.

.

.

The Setup

My daily machine is a Raspberry Pi 5.

I’ve run a couple of agents at once on the Pi before, and everything else on it slowed to a crawl. Five agents building a 3D game — plus browser checks on top — was more than I wanted to throw at it. Claude Code on the web runs on someone else’s hardware, and I could close the tab and go to sleep.

Whiteboard comparison: on the left a tiny Raspberry Pi 5 board sweating and smoking under five robot agents and a browser window; on the right a roomy cloud machine carrying the same five agents and browser with ease

But a cloud session starts empty — a bare workshop with no tools on the wall. Without a browser, Claude writes a game it can never see running. It can run tests and check that the build compiles. Whether a road actually renders on top of the terrain or disappears beneath it — that takes eyes.

cloudplay is the setup that gives every cloud session my tools — including a browser — the moment it starts.

Stay with me: more on how that changes things at the end.

.

.

.

How Sonnet 5.5 Built It

Before launching a single agent, Sonnet 5.5 wrote the task list.

Sonnet 5.5's task plan showing nine checked items — lead agent creates the architecture contract and build pipeline, five builders handle world and tracks, kart physics, gameplay and AI, art and audio, and interface, then the lead merges and tests with a browser, publishes the game, and reports back

One lead agent set up the foundation and planned how the pieces would fit together.

Five builders then worked in parallel: tracks and scenery, kart handling, race rules and AI rivals, art and sound, menus and saved progress. Once they finished, the lead merged everything and moved on to checking the result. I’ve written about both halves of this approach before — planning the build in How to Make Claude Code Actually Build What You Designed and closing the loop in How to Make Claude Code Test and Fix Its Own Work (The Ralph Loop Method).

Because Sonnet had a browser, the lead agent opened the game after integration, took screenshots, and actually looked at what was on screen.

(This is the part that makes overnight runs viable — a model that can see its own output and course-correct without waiting for you.)

It caught four visual bugs — terrain covering roads, graphics glitches, buttons falling off a small window — and routed each one back to the builder that owned that section. All four were fixed before Sonnet reported back.

Sonnet 5.5's report describing how it created tasks, wrote the architecture contract, found four bugs in its own screenshots, and sent each back to the owning agent — ending with "All four were fixed."

.

.

.

What I Woke Up To

https://youtu.be/k7oO3ZLmPwE

The first thing I did when I woke up was open the session on my phone to see if it had finished.

It had.

Kart Blitz title screen displaying the game logo with a lightning bolt, Press Start button, and the tagline An Original Arcade Kart Racer

Kart Blitz.

An original kart racer where every model, sound, and track is generated from code — no borrowed names or art.

Here’s what five agents and about five hours produced:

Track selection screen showing eight tracks across the Sprout Cup and Ember Cup — Sunny Meadow, Dune Canyon, Frostbite Peaks, Neon Skyline, Coral Cove, Haunted Hollow, Volcano Rush, and Starlight Ribbon — with Sunny Meadow selected and its description visible
  • Eight tracks across two cups, eight drivers, 50cc to 200cc
Low-poly kart on the starting grid approaching the Start banner, with the lap counter at 1 of 3, speedometer at 0 km/h, and a minimap in the corner
  • Drifting with mini-turbos, 13 items, 11 AI rivals per race
  • Grand Prix, Versus, and Time Trial with a ghost to race
  • Keyboard, gamepad, or touch controls

The comparison.

Reddit runMy run
EffortHighMax
ToolsNoneA browser to check its work
Build time67 minutesAbout 5 hours
Tracks48
Automated testsNone reported308

My run took longer and came back bigger, but the prompts differed too, so the gap reflects more than the effort setting alone.

Sonnet’s honest report.

Sonnet 5.5's honest report listing what it could not verify — game feel, frame rate, the live link, physical devices, and audio quality — followed by the line "I would not call it triple-A. It looks like a well-made low-poly arcade racer."

One line from the report stood out: “I would not call it triple-A. It looks like a well-made low-poly arcade racer.”

Here’s the thing — that kind of honesty matters more than you’d think. When a model tells you exactly where its own checking stopped, you know where to pick up in the morning.

The report becomes your starting line.

My verdict.

Sonnet undersold itself.

The one thing it couldn’t test — whether a person would actually enjoy the game — is where it surprised me.

It’s genuinely fun.

The first proper race I ran, I won: 1st of 12 at Sunny Meadow, 100cc. I expected something that would look polished in screenshots and fall apart the moment I grabbed the controls.

It held up from the first turn.

Results screen showing 1st place finish out of 12 racers at Sunny Meadow, Sprout Cup, 100cc Versus mode, with a total time of 1:21.40 and best lap of 0:26.40

I tried it on my phone too.

Mobile view mid-race with touch controls — virtual joystick on the left, Brake, Drift, Gas, and Item buttons on the right, kart at 83 km/h on lap 1 of 3

The on-screen stick and buttons worked straight away — and Sonnet had only ever tested touch with simulated taps.

Think you can beat 1:21.40 at Sunny Meadow? Play Kart Blitz here. It runs in the browser on desktop and phone.

.

.

.

Why I Called It cloudplay

So, can Sonnet 5.5 build a full, playable game from one prompt with nobody stepping in?

Yes — given the right tools and about five hours of someone else’s compute.

The name breaks down into two halves.

  • Cloud: a bigger machine that already has your tools — Firecrawl for research, a browser to check its work — the moment the session starts.
  • Play: hand Claude the repo and a prompt describing what you want, and let it play out. It builds the thing, or at the very least tests out the idea you had in mind.

Either way, you wake up with something concrete to look at.

Whiteboard diagram breaking down the name cloudplay: cloud, a bigger machine with your tools already installed, plus play, describe the idea and let it play out

That combination changes what “using Claude” looks like in practice.

Before this, my habit was hovering over the session — approving each step, nudging the direction every few minutes. (I suspect that habit sounds familiar.) This time I typed one prompt, closed the laptop, and went to bed. The browser and the tools were already in place, so I could trust the session to catch its own mistakes while I slept.

Whiteboard comparison: on the left a tired developer hovering over a laptop, approving and nudging every step; on the right the developer asleep while Claude works in the cloud

I started this before bed and reviewed it in the morning.

👉 Anything I can describe clearly enough can now run while I sleep, on a machine bigger than mine. The honest report at the end means I wake up knowing exactly what was checked and what still needs a human eye.

A kart racer turned out to be a good stress test because it demands everything at once — visuals, physics, sound, AI opponents, menus, touch controls. That overnight handoff, from one prompt to a playable game and a detailed breakdown of what worked and what didn’t, is what the whole setup was built for.

If the tools can carry a game this complex on their own, they can carry most of the ideas I’d want to try.

Whiteboard flow from evening to morning: hand Claude the cloudplay repo and one prompt, sleep while Claude builds and checks its work in a browser in the cloud, wake up to a finished game and an honest report

Try cloudplay, run your own variation of the prompt, and reply with what you built.

Nathan Onn

Freelance web developer. Since 2012 he’s built WordPress plugins, internal tools, and AI-powered apps. He writes The Art of Vibe Coding, a practical newsletter that helps indie builders ship faster with AI—calmly.

Join the Conversation

Leave a Comment

Your email address will not be published. Required fields are marked with an asterisk (*).

Enjoyed this post? Get similar insights weekly.