| ▲ | jeremyjh 4 hours ago | ||||||||||||||||
We don't have a reward function for "human understanding". We reward the appearance of understanding. We define goals that we cannot conceive of reaching without something like understanding happening. There is something happening, but it is alien and counter-intuitive - it makes bizarre mistakes that betray it - and we don't know what it is. I'm pretty sure it is not human understanding. | |||||||||||||||||
| ▲ | antonvs an hour ago | parent [-] | ||||||||||||||||
> We reward the appearance of understanding. Which is exactly what happens with human evolution and development. Sure, we can say LLMs don’t have “human” understanding - which is something we can’t really define anyway - as long as we’re not trying to claim LLMs don’t have understanding at all. The latter is a much higher bar. > We define goals that we cannot conceive of reaching without something like understanding happening. Functionally speaking, that is understanding. Again if you want to go past a functional definition, that’s a bar which no one can clear right now. | |||||||||||||||||
| |||||||||||||||||