logoalt Hacker News

altcognitolast Monday at 4:53 PM1 replyview on HN

Well, yes, but remember there is the reinforcement learning that is applied after, and the system prompts that will bend the results.


Replies

lossololast Monday at 5:12 PM

Yeah, agree on both points. You can embed any bias you want using RL, regardless of training data.