Remix.run Logo
janalsncm 4 hours ago

I don’t think it’s misleading if you compare on the use cases they suggested. It’s faster and cheaper (no idea if higher quality), so it’s immediately interesting for certain things.

And if you buy their RLCD claims, this might be even better than huge models that know a bunch of irrelevant things.

WhitneyLand 4 hours ago | parent [-]

What was misleading was the original title:

"Jev: New frontier model 40-400x cheaper and 20-200x faster"

I'm not the gatekeeper of who gets to call themselves a frontier model, but I don't think most people would count Jev in that group. It sounds false.

If their specific claims hold up, then it would make more sense to say something like:

"Advanced the speed/cost frontier for structured decisions"

sroussey 3 hours ago | parent | next [-]

I dunno, I would consider Waymo and Tesla to have frontier models.

I think AlphaFold and related are also frontier models.

Being an LLM does not seem like the qualifier for frontier.

riknos314 an hour ago | parent | next [-]

This is likely still an LLM (in the purest definition of a language model with relatively many parameters) since the inputs are natural language, just not a generative LLM as the output is something other than more language.

2 hours ago | parent | prev [-]
[deleted]
bigglebear 20 minutes ago | parent | prev | next [-]

They're making it sound as if it's a frontier LLM (on purpose), while they cut out all of the intelligence that autoregressive token generation gives you.

janalsncm 3 hours ago | parent | prev | next [-]

Large language models are not the only type of model.

alfalfasprout 3 hours ago | parent | prev | next [-]

How is this not a frontier model? It's bleeding edge in its own niche. It's not a frontier LLM; however, applicable to many of the things people use LLMs for.

bigglebear 32 minutes ago | parent [-]

It's nothing like a traditional LLM and so should not be compared to one. It's a heavily constrained, tiny model that can only produce a probability score or a yes/no answer over pre-defined selections. It has no long-context capacity.

I mean, imagine comparing this thing to Astra, it's hilarious. They don't even tell you what the max input size is, and they only allow 10 possible answers to choose from for the Choice mode. It's probably like a 1billion param model. They say it's "not small", but there's zero reason to believe that.

I suspect someone will be able to recreate this within a week by piecing together open-weight models.

nalishwana 2 hours ago | parent | prev [-]

nali shwana