Twitter/X

On 2026-06-25, Opus 4.8 pushed Liquid AI’s LFM2.5 230M to 1,400 tok/s running…

Brief

Xenova (@xenovacom) reports their agentic WebGPU kernel optimization framework continued after Fable 5’s shutdown and, on 2026-06-25, used Opus 4.8 to push Liquid AI’s LFM2.5 230M to 1,400 tok/s running locally in the browser. They released the demo and kernels for verification and noted Fable earlier reached 255 tok/s on Gemma 4.

Why it matters

On 2026-06-25, Opus 4.8 pushed Liquid AI’s LFM2.5 230M to 1,400 tok/s running locally in the browser; Xenova released the demo and the generated kernels for users to run themselves.

Key details

  • Before Fable 5 was shut down, Fable reportedly pushed Gemma 4 to 255 tok/s on WebGPU; Xenova reiterates that claim and provides the kernels for verification.
  • Xenova presents an agentic WebGPU kernel optimization framework that kept running after Fable 5’s shutdown and calls agentic kernel optimization 'the future of on-device inference.'
Source evidence

While we eagerly await Fable 5's return, our agentic WebGPU kernel optimization framework kept running.

Opus 4.8 picked up where Fable left off, pushing Liquid AI's new LFM2.5 230M to an unbelievable 1,400 tok/s... running locally in your browser.

Don't blink or you'll miss it.

Video

Xenova (@xenovacom)

Before Fable 5 was shut down, it pushed Gemma 4 to 255 tok/s on WebGPU. Some didn't believe it was real.

Today we're releasing the demo and kernels it wrote for you to see yourself. Run it locally in your browser.

Agentic kernel optimization is the future of on-device inference

Video

— https://nitter.net/xenovacom/status/2067289897111638484#m