Twitter/X

GPT-5.6 Sol achieves state-of-the-art (SoTA) performance on the ARC-AGI-3…

Brief

Tibo reports that GPT-5.6 Sol reaches SoTA on the ARC-AGI-3 benchmark after making two configuration changes: allowing the model to reason across multiple context windows and employing OpenAI's canonical compaction implementation. The brief post (dated 2026-07-30) links to OpenAI documentation showing how the two settings produce the improvement.

Why it matters

GPT-5.6 Sol achieves state-of-the-art (SoTA) performance on the ARC-AGI-3 benchmark according to Tibo (@thsottiaux).

Key details

  • Improvement required only two setting changes: enabling multi-context-window reasoning and using OpenAI's canonical compaction implementation (link: openai.com/index/how-two-set…).
  • Claim posted on 2026-07-30 and shared via a short 'goblin-level blog post' style note linking to a demonstration/guide.
Source evidence

goblin-level blog post

Tibo (@thsottiaux)

Turns out GPT-5.6 Sol is actually SoTA on ARC-AGI-3.

Just took two setting changes. You just have to allow it to reason and work over multiple context windows with the help of our canonical compaction implementation.

openai.com/index/how-two-set…

— https://nitter.net/thsottiaux/status/2082609662231502932#m