They don't lead on latency nor quality; but their execution was superb
Who they trailing on quality?
Here's another leaderboard: https://huggingface.co/spaces/multimodalart/jev-decision-ind...
I'm not the person you're asking, but:
https://benchmarkheaven.com/jev-models
According to this benchmark, Jev is currently trailing Quyet-1.0-Large and a few other hastily put-together LLM-based decision API-like setups.
And the 'better' ones are slower and cost more for a tiny bit more accuracy. Its hard to sell that as being better when speed and price have been JEVs main selling points.
What even is this? What does it do?