| ▲ | siva7 6 hours ago | ||||||||||||||||||||||
Sounds fun. As fun as their press release claiming it is the most safety aligned model ever. | |||||||||||||||||||||||
| ▲ | isoprophlex 5 hours ago | parent | next [-] | ||||||||||||||||||||||
It's super aligned! It can hide its thoughts! There is no evidence of steganographic thought masking, there is nothing to worry about! It has become better at cheating! Maybe they don't know themselves what's really going on. We are all in the interesting times gang now. | |||||||||||||||||||||||
| ▲ | paxys 5 hours ago | parent | prev | next [-] | ||||||||||||||||||||||
The model said it was perfectly aligned. | |||||||||||||||||||||||
| |||||||||||||||||||||||
| ▲ | NBJack 5 hours ago | parent | prev | next [-] | ||||||||||||||||||||||
Hey, don't forget how "dangerous" GPT-2 was supposed to be. | |||||||||||||||||||||||
| |||||||||||||||||||||||
| ▲ | 6gvONxR4sf7o 5 hours ago | parent | prev | next [-] | ||||||||||||||||||||||
So, probably most aligned as measured by the metrics that are the least reliable on it. | |||||||||||||||||||||||
| ▲ | wilg 4 hours ago | parent | prev [-] | ||||||||||||||||||||||
These are not mutually exclusive ideas | |||||||||||||||||||||||