logoalt Hacker News

applfanboysbgontoday at 10:08 PM0 repliesview on HN

> much more prose than code to train on

This is completely irrelevant. An LLM's voice is not influenced by its at-scale training data. LLM-prose is the result of human feedback training rewarding it for making every single sentence sound like a YouTube or BuzzFeed headline, which is effective against a large portion of the population who get dopamine hits from such clickbait style of writing, and completely infuriating to the portion who recognize it for what it is.

LLM code is also infuriating, but for a different reason, because it is not reinforced in exactly the same way. However, people are fine with it because they don't read it. If they actually read the dogshit code an LLM produces, and were capable of understanding it, they would be infuriated by it too.