logoalt Hacker News

epolanskiyesterday at 11:50 AM1 replyview on HN

It's more complex than that, especially as post training is often goal based.


Replies

OscarCunninghamyesterday at 12:39 PM

I wouldn't have expected that there was post training specifically on the issue of looking for proofs vs counter examples. But it might be that other post training has a side effect of making AIs better at looking for counter examples. I wonder if these agents are overall less biased and more rational than humans. Can you expand on what you mean by goal based training?