| ▲ | jakozaur 4 hours ago | |||||||||||||
Yeah, real Jev got really weird, no benchmarking clause. Their Terms of Use (1(v)) and MCA (2.3(f)) both prohibit users from publishing "benchmarks or performance information about the Services". No major AI has it; we are back to Oracle-style legal. Though Jev is original, it looks highly replicable. | ||||||||||||||
| ▲ | sodimel 4 hours ago | parent | next [-] | |||||||||||||
I'm working on something from a crappy laptop, those numbers from jev can totally be matched: | ||||||||||||||
| ||||||||||||||
| ▲ | cmrdporcupine 3 hours ago | parent | prev [-] | |||||||||||||
There's also prior art. Or probably, anyways. https://www.reddit.com/r/LocalLLaMA/comments/1wijo3e/i_liter... Not only is it replicable as you say, things like it already exist(ed). The important bit of course is in the actual implementation: a) models fine tuned to produce good results for these types of questions and b) runtimes optimized to do this quickly and at scale | ||||||||||||||