Build 01 Best overall product
GPT 6 Astra
The most complete interpretation
A refined mission desk, clearer code and review workflow, a model library, and strong defaults. It produced the best overall product, although the improvement over GPT 5.6 Sol was incremental.
- Effort
- Similar elapsed time to GPT 5.6 Sol
- Stack
- React · TypeScript · Monaco · Tauri 2
Explore this build on GitHub ↗ Build 02 Efficiency winner
GPT 5.6 Sol
The best fit for this prompt
A polished mission-based workspace with editing, Git diffs, terminal access, model context, and visible agent actions. It needed very little follow-up and delivered the strongest balance of quality and efficiency.
- Effort
- Two small follow-up turns
- Stack
- React · TypeScript · Monaco · Tauri 2
Explore this build on GitHub ↗ Build 03 Most ambitious workflow
Fable 5.1
Capable, elaborate, and expensive
Threads, acceptance criteria, a Sharpen action, model roles, project isolation, and replay made this the most opinionated workflow. It worked, but required UX fixes and heavy token use to reach the finish line.
- Effort
- About 14 hours including usage-limit waits
- Stack
- React · TypeScript · CodeMirror · Tauri 2
Explore this build on GitHub ↗ Build 04 Speed standout
DeepSeek V4 Pro
Lightning fast with major gaps
It assembled a recognizable IDE with files, an agent view, models, and source control in roughly 14 minutes. Missing Git diff inspection and other functional gaps kept the result from being a usable daily tool.
- Effort
- Roughly 14 minutes
- Stack
- React · TypeScript · Monaco · Tauri 2
Explore this build on GitHub ↗ Build 05 Did not hold together
GLM 5.3
Opened, but core workflows stayed broken
The app could open and load files, but key interactions still failed after three attempts to repair it. The run is included because failed products are part of the experiment, not something to edit out of the story.
- Effort
- Three unsuccessful repair turns
- Stack
- Svelte · TypeScript · CodeMirror · Tauri 2
Explore this build on GitHub ↗