| ▲ | thesz an hour ago | |
In the training of LLMs, the order of batches and their content can introduce adversarial behavior [1].[1] https://www.pure.ed.ac.uk/ws/portalfiles/portal/256761768/Ma... You can get all the source to validate training and weights and still end up with adversarial behavior in the model. | ||