I guess it’s only my opinion but having used grok for personal chat: it’s by far the worst one amongst Claude, ChatGPT and even Deepseek, Gemini etc.
The personality is bland and it doesn’t work nearly as hard or even tries to help.
> it doesn’t work nearly as hard
Until you ask it to start generating horrific imagery and then it's best in class.
I used openrouter to send same prompt to qwen, derpseek, gemini and grok and found that grok does good research and produces less bullshit, especially when prompted to be critical of an idea
> The personality is bland
Sounds like a plus. Guess I will give Grok another try...
This has been my experience as well. Grok will end tasks almost immediately and claim "Done!". It's definitely the laziest and most "dishonest" of all the models. The others aren't perfect, but I can't use Grok for any serious coding task.
> The personality is bland
I don't use Grok, but do you want your LLM to have a personality? "Personality" is exactly what people don't like about Claude.