COMPUTER VISION JUST TOOK A MASSIVE LEAP FORWARD
A few hours ago @perceptroninc released Agentic Detection, a brand new API that can localize objects just by describing them or showing a crop 🤯
Standard one-shot detectors do a single pass and miss critical details.
Perceptron fixes this with an agentic harness that actually takes control of the image.
The Mk1 model zooms, crops, and tiles completely on its own.
It issues multiple calls per request to catch tiny, hidden objects that others miss!
You can use the API in 3 ways:
▪ Detect everything → an exhaustive inventory with zero labels required.
▪ Open-vocab categories → Find anything using text prompts.
▪ Visual exemplars → Find items matching a single visual crop!
This handles satellite imagery, robot sensors, LiDAR, and more.
And at just $0.15/M input and $1.50/M output, the pricing is highly disruptive.
Check out these 5 wild demos below 🧵↓
Video