| ▲ | embedding-shape a day ago | ||||||||||||||||
> What is needed the most right now is something similar to Bonsai 27B, with a modest memory footpint, but faster and more capable Yeah, that'd be neat, but that's not what this announcement is about at all: > With a massive 2.4T parameters | |||||||||||||||||
| ▲ | docheinestages a day ago | parent | next [-] | ||||||||||||||||
True. It was more of an open letter, with hopes that the Qwen team sees the comments in this thread. | |||||||||||||||||
| ▲ | cyanydeez a day ago | parent | prev [-] | ||||||||||||||||
dont we all deem the ability to improve large models as the defacto capability to produce small ones? | |||||||||||||||||
| |||||||||||||||||