Now the DwarfStar laguna-s2.1 branch contains a GPT5.6-Sol coded (and manually tested for apparent sanity with non trivial tasks) Laguna S2.1 implementation. 50 t/s generation, 500 t/s prefill on M5 Max. Only Metal for now. GPT5.6 Sol can write for you the other backends.
DwarfStar 'laguna-s2.1' branch now includes a Laguna S2.1 implementation coded by…
Brief
The DwarfStar repository's laguna-s2.1 branch hosts a GPT5.6‑Sol–written Laguna S2.1 implementation that the author says was manually sanity‑tested on nontrivial tasks. Measured speeds on an M5 Max are about 50 t/s for generation and 500 t/s for prefill; only the Metal backend is available so far, but GPT5.6‑Sol is claimed able to generate additional backends.
Why it matters
DwarfStar 'laguna-s2.1' branch now includes a Laguna S2.1 implementation coded by GPT5.6‑Sol and manually tested for apparent sanity on nontrivial tasks (author: @antirez).
Key details
- Performance reported on an Apple M5 Max: ~50 tokens/second generation and ~500 tokens/second prefill; implementation currently supports only Metal backend, with GPT5.6‑Sol claimed able to produce other backends.