i found the first model that's physically incapable of hiding why it hallucinated...
ChatGPT, Claude Gemini... every model you use is a black box
it gives you an answer, you have no idea where it came from, and when it's wrong you can't tell why - you just reroll the prompt and hope
but this might be the first AI you can actually argue with about why it said what it said lol
these guys just flipped that open
it runs on a model built so the reasoning is readable from the start, not bolted on afterward
ask it something and you get the usual answer, plus a panel showing:
> the actual concepts it used to think (wildlife, taxonomy, etc.)
> which exact chunks of training data the answer came from
so when it says something weird, you simply click the sentence and see what was driving it
that's the difference between a tool you prompt and a tool you can audit
Guide Labs (@guidelabsai)
The first inherently interpretable AI platform is finally here. Welcome to Clarity.
Video
— https://nitter.net/guidelabsai/status/2062554858092449961#m