Twitter/X

Impossible Research introduced an agent harness named [schema] that they say…

Brief

Impossible Research’s team (announcement credited to Haven Feng) released [schema], an agent harness that they claim enables LLMs to play games, write code, and reason like a physicist. They report 99% RHAE with Opus 4.8 + Fable 5 and 95.35% RHAE with GPT-5.6 Sol on the ARC-AGI-3 Public set, and provided a video demo.

Why it matters

Impossible Research introduced an agent harness named [schema] that they say plays games, writes code, reasons like a physicist, and saturates the ARC-AGI-3 benchmark.

Key details

  • [schema] achieved 99% RHAE on the ARC-AGI-3 Public set using Opus 4.8 + Fable 5, and 95.35% RHAE using GPT-5.6 Sol, per Haven Feng.
  • The announcement frames [schema] as a method to make an LLM “think like a physicist” and includes a video demo from Haven Feng.
Source evidence

From our team at Impossible Research - an agent harness that plays games, writes code, reasons like a physicist, and saturates the ARC-AGI-3 benchmark.

Haven Feng (@HavenFeng)

Today, we’re introducing [schema]: a harness reaching 99% RHAE with Opus 4.8 + Fable 5 and 95.35% with GPT-5.6 Sol on ARC-AGI-3 Public set.

[schema] makes an LLM think like a physicist. 🧵

Video

— https://nitter.net/HavenFeng/status/2077770348876247502#m