Twitter/X

Ori Eval launched on 2026-08-03 by OpenRouter (announcement shared by…

Brief

OpenRouter released Ori Eval (announced 2026-08-03) as a lightweight evaluation tool that executes model calls through OpenRouter’s APIs against tasks in your codebase and scores the results. The project positions itself as an easy entry point for writing evals and emphasizes selecting the best model per task rather than claiming a single universally best model.

Why it matters

Ori Eval launched on 2026-08-03 by OpenRouter (announcement shared by @alexatallah) as a tool to make it “the easiest way to write your first eval.”

Key details

  • Ori Eval uses OpenRouter’s APIs to run and evaluate model outputs per task in your codebase, emphasizing there’s no single best model — only the best model for each task (example install shown via curl to openrouter.ai/skills/spawn-o…).
Source evidence

It's now launched! Try out Ori Eval here:

OpenRouter (@OpenRouter)

Introducing Ori Eval: the easiest way to write your first eval.

There's no definitive best model, only the best model for each task. Ori Eval leverages OpenRouter's APIs for each task in your codebase, and then evaluates the results.

curl -fsSL openrouter.ai/skills/spawn-o…

— https://nitter.net/OpenRouter/status/2084301100078027143#m