Introducing EvalRouter by @KimptonAI.
One API for evaluating any model against any benchmark.
We built it because running an evaluation still means setting up a harness, provisioning sandboxes and babysitting long jobs.
With EvalRouter you name a model and a benchmark and get