every scenario runs against a real Supabase environment. Agents hit our MCP server and CLI and then are scored with deterministic checks plus an LLM judge
Read the blog → supabase.com/blog/introducin…
Link
Introducing Supabase Evals
Our open-source benchmark for how well AI coding agents build with Supabase.
supabase.com