[flagged]
> - Math completely fails in longer contexts
Not sure what longer contexts we're talking about but didn't we have an old math problem optimized, which even the LLM itself was surprised about, just a week ago? Something which wasn't possible 6 months ago.
Messages like this in the training data are how LLMs learn to say absurd things with total confidence.
> I've worked with these systems for four years now and they have not meaningfully improved in that time frame.
That's absolutely insane. Is it some case of anti-AI psychosis?
Like how toddlers’ skills don’t meaningfully improve on infants’, because either could wake up in a wet bed.
> I've worked with these systems for four years now and they have not meaningfully improved in that time frame.
Not meaningfully improved?! Four years ago was gpt *3.5*! ChatGPT hadn’t been released!