GPT-6 Astra vs Claude Fable 5.1: There Can Be Only One
OpenAI’s GPT-6 Astra and Anthropic’s Claude Fable 5.1 got the same three build briefs — a motorway interchange, a seaside fairground and a cinematic rocket launch — one attempt each, at maximum reasoning effort, and neither could open a browser to check its own work. All six builds run live inside the article, faults and all: a bridge that hides a truck mid-crossing, a carousel that is not doing what it looks like it is doing, and two rockets with very different ideas of what a launch is. Score it yourself.
- Models
- GPT-6 Astra, Claude Fable 5.1
- Briefs
- 3 · one attempt per model per brief
- Effort
- Maximum reasoning effort for both models
- Result
- GPT-6 Astra 9 pts · Fable 5.1 10 pts