logoalt Hacker News

whstltoday at 3:54 PM0 repliesview on HN

> the behavior is meant to have you shut up and trust it and not ask questions

This seems to be exactly the kind of thing automated/massive training would produce, just like it did with sycophancy recently.

Claude users would just gave up after the word vomit and some classifier considered it a success and into the model it went.

Wrong incentive and nobody checking.