The live demo got too good to leave on-screen: play the 3D Battleship game from this video → https://battleship.metal-sole.com
GPT-5.6 is a great model. But the more interesting shift is what happens when we stop handing AI a list of tasks and give it an objective instead. I put Sol, Terra, and Luna through my benchmark—then challenged GPT-5.6 Sol and Fable 5 to build a cinematic 3D Battleship game from one high-level brief. The live result is exactly why “better, faster, cheaper” is no longer a big enough story.
I’ll show the three new GPT-5.6 models, what I saw on the CARE benchmark, and the real build runs behind the verdict. Fable produced the more polished Battleship build; GPT-5.6 Sol finished roughly three times faster. The bigger takeaway is not a single winner—it is that models at this level can increasingly turn an objective into the work required to reach it.
Chapters
00:00 GPT-5.6 and the next AI inflection point
01:27 The live Fable vs. GPT-5.6 build-off begins
02:03 Sol, Terra, and Luna
03:14 The CARE benchmark
05:30 The live Battleship builds are running
06:35 The native-app build comparison
09:50 What /goal changes
14:17 Fable 5's Battleship build
15:31 GPT-5.6 Sol's Battleship build
17:46 The four AI inflection points
20:38 From task-level to objective-level work
23:00 Fable vs. GPT-5.6: the verdict
GPT56 #OpenAI #Codex #CodexCLI
Channel: Matt Maher
Published: 2026-07-13
Video URL: https://www.youtube.com/watch?v=4tCIa_fnEIo