| ▲ | adam_rida 3 hours ago | |
The evaluator is public here: https://echo.tracerml.ai/eval/ It currently exposes 907 stored rows across seven benchmark families, with prompts, outputs, grades, and cost records. More benchmarks are coming soon. Echo does not disclose its per-request routing decision because that policy is the product. We can, however, publish some of the eligible open-weight model pool, version dates, aggregate allocation mix, and evaluation settings without exposing the request-level recipe. New video is also being made. | ||
| ▲ | seizethecheese 40 minutes ago | parent [-] | |
Isn't the relevant benchmark RouterBench? https://arxiv.org/html/2403.12031v2#S7 | ||