| ▲ | 8note an hour ago | |
does lora do a good job at teaching the model new things that werent in the training data? without trillions of examples of following instructions at a million context length, im not convinced the behaviour is in the weights to begin with | ||