logoalt Hacker News

winwangyesterday at 8:12 PM1 replyview on HN

I found 4.6 more amenable than 4.8 to style directions, we'll see how 5.0 does. Super-small-sample-size: I think part of its "Claude-ism" style comes from its propensity to try and "proactively" move the conversation/work along. Not sure how this would fare in non-obviously-productive environments, I'd guess "it's still annoying" considering your evidence.

I'm also thinking of another benchmark: (quantified) stylistic range across different prompts. Just putting it out there if anyone wants to do the work for me :D


Replies

sibeliusstoday at 1:27 AM

4.6 is night and day better. It was before the big language switch up. Terrible direction that Anthropic has taken this.