logoalt Hacker News

lukeschlatheryesterday at 10:34 PM0 repliesview on HN

I don't think the problem is a lack of previous research, the problem is it looks like the target metrics were selected by an LLM, and it's unclear what the LLM was told to optimize or if it was just told to try and make a better KV-cache.

What I've noticed with Claude is that regardless of how I prompt, it will find a few metrics to optimize. Often the metrics it chooses to optimize have zero relation to the actual metrics I want to optimize, which are ones that cannot be measured without more work than Claude can do in a single 1M token context window. It's really hard to stop Claude from optimizing whatever metrics it can find when the actual metrics I want to optimize are not computable.