| ▲ | overfeed 4 days ago | |
I'm just glad the dweebs who were parroting "OMG, Distillation attacks!!1!" have been empirically proven wrong, and discussion around open-weight models are more rational now. | ||
| ▲ | happycube 3 days ago | parent [-] | |
Distilled data or no, the Chinese labs are bringing a lot of solid original work in training and inference efficiency. Deepseek's work has probably made everyone's AI cheaper to run by this point. | ||