prime intellects new model harness is basically worth 3x the model upgrade itself which is fckin nuts
5 weeks ago the best ai model scored <1% on ARC AGI-3.
1.5 weeks ago opus 5 scored 30% (#1 at the time)
today that same opus 5 model scored 95.5% beating human experts. 60+ point jump
all while using fewer tokens!
you get a smarter model for a much cheaper rate. why wouldn’t you use this?
(open source btw)
Prime Intellect (@PrimeIntellect)
Prime Agent is a general-purpose coding harness
On ARC-AGI-3, it scores 95.5%, surpassing the human-expert baseline, but the gain is not benchmark-specific.
We see major improvements across models when compared to their proprietary harnesses:
— https://nitter.net/PrimeIntellect/status/2085087000764568010#m