We ran a test between Fable 5, GPT-5.6 Sol, and Kimi K3. One prompt. Across all three models. With our /design command. Results: We reviewed gameplay, UX, and UI
- Kimi K3 performed well across all three.
- Fable 5 is good, but its UI doesn’t look much better.
- GPT-5.6 Sol has the best UI, but the gameplay is worse.
Ranking (DX, features, and cost):
- Kimi K3: 9.5/10 · $0.12
- Fable 5: 9/10 · $0.60
- GPT-5.6 Sol: 7/10 · $0.32
Our engineering and design team has been testing 20+ side-by-side comparisons across frontier and open models. All runs are public and open source. Benchmark for this demo here: github.com/CommandCodeAI/slash-design-showcase/tree/main/subway-surfers
Try Command Code
The best coding agent for open models, in your terminal or on your desktop.
1npm i -g command-codeDesktop App · Docs · X · Discord
