| ▲ | CamperBob2 an hour ago | |||||||
Suggesting otherwise has become an extraordinary claim requiring extraordinary evidence. VibeThinker 3B constitutes extraordinary evidence, IMO. The first such evidence I've seen myself. Very small model, very low literacy, almost no world knowledge, but it is as good at math and logical reasoning as models a hundred times larger. The Bitter Lesson is a valid and trenchant observation about how about we got here, but I think it's a mistake to assume it tells us very much about where we're going. Too much has changed recently and is still doing so. | ||||||||
| ▲ | spockz 11 minutes ago | parent | next [-] | |||||||
So theoretically, if you give that model the means to find information, ascertain the quality of said information, it could still reason its way to an proper answer? Is this whole thing than maybe a read vs write optimisation again? Spent more time and effort training more knowledge into the model upfront and get it out in a single question instead of training a small model and needing more steps to answer the same question? | ||||||||
| ▲ | algo_trader 15 minutes ago | parent | prev [-] | |||||||
> VibeThinker 3B constitutes extraordinary evidence.. math and logical reasoning Any similar model aimed at coding? A >10B model for mass spawning/swarming and reporting back to a larger model | ||||||||
| ||||||||