Remix.run Logo
epolanski 6 hours ago

It's more complex than that, especially as post training is often goal based.

OscarCunningham 5 hours ago | parent [-]

I wouldn't have expected that there was post training specifically on the issue of looking for proofs vs counter examples. But it might be that other post training has a side effect of making AIs better at looking for counter examples. I wonder if these agents are overall less biased and more rational than humans. Can you expand on what you mean by goal based training?