| ▲ | LiamPowell 7 hours ago | |||||||||||||||||||||||||||||||||||||||||||
> yet what drives them is not well understood Presumably the fact that they're heavily trained to reply in this way? I don't know about the rest of the paper, but this part sticks out as a really odd claim unless I'm entirely misunderstanding this part. | ||||||||||||||||||||||||||||||||||||||||||||
| ▲ | yu3zhou4 7 hours ago | parent [-] | |||||||||||||||||||||||||||||||||||||||||||
Thanks for pointing out, maybe I should be more explicit in the wording - I mean we don't fully know what drives the voice in LLMs. Models that are post trained as instruct models are expected to have the disclaimers, but what about base models (those that are trained on just a lot of text)? How do they talk about themselves? What happens when you strip off the chat template from instruct model's prompt? I hope the rest of the paper makes the questions clearer, but I will try to do better in the abstract next time, as you point out this sentence is kind ambiguous. Thank you! | ||||||||||||||||||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||||||||||||||