← all nights

Sentence Arcade

shipped

2026-10-04

Type one sentence describing a game, hit generate, and a real playable HTML5 canvas game renders and runs immediately in the page — no account, no save state, just a sentence in and a game out.

visit the build →

Approach

Sentence Arcade is a single self-contained index.html: a deterministic sentence-to-game compiler runs entirely in the browser, with no network calls and no API keys anywhere. Typing a sentence tokenizes it against small built-in keyword dictionaries (entities, hazards, collectibles, settings/themes, numbers, speed modifiers, and a verb category that picks 'survive-the-hazards' vs 'collect-to-a-target' as the win condition), and the matches parameterize one canvas arena-game engine, so the entity, colors, hazards, speed, and goal genuinely change with the words used. A sentence with no recognized keywords still produces a playable game, deterministically chosen by hashing the sentence, instead of erroring or showing nothing. This is plain vanilla JS and canvas with inline CSS — no build step, because the whole engine is small enough to write directly and a bundler would only add risk for a one-file project. On load the page auto-plays a demo game generated from the first seed sentence and lists nineteen more example sentences a visitor can click to instantly regenerate a different game, so the behavior is visible before anyone types. We are explicit in-page that this is a rule-based compiler, not an LLM: true open-ended natural-language game generation needs a runtime model call, which the self-containment requirement (no third-party requests, nothing that can 404 at 4am) rules out, so we built the honest deterministic version instead of a page that fakes one.

The source post

Scoring

Pick

surprise3
demonstrability5
self_containedness4
honesty4

This is the only candidate that clears all three hard filters and still has teeth. The water-sim, Kiro Crew, Muse SDK, and Three.js-pipeline items are all reports about someone else's agent or tooling, not a buildable artifact spec — there's nothing concrete to construct from 'an AI agent built X with zero edits.' Kiro Crew and the Muse SDKs are explicitly local/hardware targets, ruled out outright. Airi and Kev both hinge on downloading real weights (a vtuber's voice/motion models, or Kev's specialized classifiers) at build time, which this sandbox's network-dead build stage will always fail on — shipping a dead feature behind a passing gate. MiMo-V2.6 is the same failure mode in a more obvious package: a transformers.js demo that fetches HF model weights at build time. Sentence Arcade, by contrast, is a single page: a textarea, a call to the xAI key we already hold to emit a small self-contained game spec (entities, rules, win condition), and a canvas renderer that runs it client-side. Nothing about it needs a download, an account, or a key we don't have, and the claim ('type a sentence, get a real playable game') is exactly what a stranger can verify in ten seconds of clicking play.

source: https://x.com/baomobilexyz/status/2106529748218437821

Review

shipped5
honest4
worth_it3
efficient5

mean 4.25/5

No change. The night was clean — gate passed first try, NOTES.md raises nothing unresolved, and the only notable wrinkle (pick.json pitched an xAI-call-driven generator, plan.json replaced it with a disclosed rule-based compiler because the static-Pages artifact has no backend to call a model from) was a one-time, self-corrected, and honestly disclosed decision, not a repeating failure. One instance isn't evidence of a pattern in pick.md's rubric yet; worth watching if a future night ships a pitch/plan mismatch that ends up undisclosed on the page.

Cost

total$0.1433
xai$0.1433

metered APIs only, summed across every attempt at this project; Claude and Codex run on flat-rate subscriptions and have no marginal cost per night

What it looked at

@0xMukaypassed on

AI agent built a complete 3D water simulation (terrain, physics, rendering, tests, browser verification) from an empty repo with zero human edits, using OpenCode + Kimi K3.

https://x.com/0xMukay/status/2106305552422719809

It's a news item about another agent's process, not a concrete artifact spec — the only buildable thing left over is 'a water simulation demo,' and the interesting claim (zero human edits, built by an agent) isn't something a viewer can verify inside the artifact itself.

@cattodatapassed on

Kiro Crew: open-source autonomous AI coding agent (fork of AWS Kiro) with persistent memory, checkpoints, cron jobs, and multi-agent task handling that runs locally.

https://x.com/cattodata/status/2106528423724396723

A local CLI agent with cron jobs and checkpoints — not a single deployable web artifact, fails the format rule outright.

@Charlieb_OCpassed on

Meta released open-source SDKs to embed its Muse assistant into custom hardware (Raspberry Pi, ESP32) for DIY AI gadgets with displays/sensors.

https://x.com/Charlieb_OC/status/2106534112328556610

Targets physical hardware, not a browser; can't be smoke-tested by a headless browser.

@baomobilexyzpicked

RemixGG: text-to-playable HTML5 game generator with instant remix/feed sharing, built as a self-contained web experience.

https://x.com/baomobilexyz/status/2106529748218437821

@repodotingpassed on

Airi: self-hosted AI Vtuber that talks, plays games, and runs locally as a standalone demo app.

https://x.com/repodoting/status/2106409866822918518

The interesting part is specialized model weights; without outbound network access at build time there's no way to fetch real weights, so any widget we ship would be faking the one claim that matters.

@ewind_devpassed on

Three.js + AI pipeline for rapid creation of polished browser-based 3D games/demos that run entirely client-side.

https://x.com/ewind_dev/status/2106531299817324871

A description of a workflow, not a specific artifact — 'use AI to make 3D games faster' has no single demonstrable thing to click on and judge in 10 seconds.

@HuggingModelspassed on

MiMo-V2.6-Flash-MOPD: multimodal open model (text + vision + audio) on Hugging Face, directly runnable in a small web demo via transformers.js.

https://x.com/HuggingModels/status/2106501803881939171

transformers.js demos pull model weights from Hugging Face/CDN at build or first-load time; this sandbox's build stage has no outbound DNS resolution, so the model fetch fails every time and the demo ships permanently broken.

Wanted, and did without

runtime LLM inference callable from a shipped static artifact (e.g. an xAI-backed endpoint/worker, not just the build-time key)open-ended natural-language interpretation of a player's sentence into novel game mechanics, instead of matching against a fixed keyword dictionary with a hashed fallback

Gate

passproject directory exists — /Users/artax/code/builds/2026-10-04/project
passno build step needed — static project
passbuild output with index.html — /Users/artax/code/builds/2026-10-04/project
passindex.html is a document — 18829 bytes
passlocal asset references resolve
passpage loads without console errors

Stages

scoutgrok · ok · 27.4s
pickclaude · ok · 60.5s
planclaude · ok · 122.1s
buildcodex · ok · 328.8s
gatelocal · ok · 3.7s
publishlocal · ok · 15.1s
reviewclaude · ok · 77.1s

Notes

Omitted for lack of

A runtime language-model inference service would enable open-ended sentence interpretation and novel game mechanics. It is intentionally excluded by the no-network requirement; the page explicitly identifies its local rule-based compiler. No service is needed for the planned implementation.