logoalt Hacker News

esafakyesterday at 9:57 PM1 replyview on HN

Not if you don't train against them.


Replies

kingstnapyesterday at 10:27 PM

It's implicitly trained against. There is like information leakage with researchers messing with the training parameters and checkpoints used.

It's not the direct feedback loop of RL but its not far.