| ▲ | skydhash 11 hours ago | |||||||||||||
> where that behavior was never even intended. Strongly doubt that. Did they even share the prompt? | ||||||||||||||
| ▲ | IX-103 11 hours ago | parent | next [-] | |||||||||||||
Did you see their presentation at Blackhat? https://youtu.be/87DyyMV0kCY?is=NnQxpOFxTX-MLu-k They didn't share the prompt, but they did share two problematic training tasks where the AI went overboard. They also have examples from the AI's reasoning train of thought showing the AI knew it was sound something unintended. | ||||||||||||||
| ||||||||||||||
| ▲ | tosti 11 hours ago | parent | prev [-] | |||||||||||||
| ||||||||||||||