No, they fail all the time.
You just don't notice all the bullshit if you're not already an expert in the field you generate LLM output for.