| ▲ | ngruhn an hour ago | |||||||
But maybe you can instill properties like shame during training. Models sometimes blatantly lie and cheat. In a social context, where actors remember, that might work the first time but you get penalized in subsequent tasks with loss of trust. | ||||||||
| ▲ | altmanaltman 33 minutes ago | parent | next [-] | |||||||
How do you "install properties like shame"? How is that even possible? Shame is a reaction driven by feelings and our inner selves. A model "feeling shame" is just a representation (false) and not an expression (true). Thinking that models "lie and cheat" is the first mistake since they are not consious agents who have any free will or consiousness. They do not (no matter what Dario says). Shame will just be another if-then rule if you implement it this way and will not work. Its like asking a rock to feel sad about being a rock. It literally cannot. | ||||||||
| ||||||||
| ▲ | xscott 36 minutes ago | parent | prev [-] | |||||||
[dead] | ||||||||