I don't trust agents writing specs without human approval.
I've been bitten by agent-written ADRs. Agents carelessly add extrapolated details and speculated nice-to-haves that I never asked for, and this becomes a source of bloat that keeps coming back like a boomerang.
Perhaps this is precisely the type of work product that humans should produce instead of LLMs. Humans are the ones making these decisions, after all.
Of course I meant you review the docs the agent writes (as well as everything else the agent writes)
I’ve found that Agents cannot differentiate between a general principle or abstraction to be followed vs a one-off code review comment that’s applicable only to a single PR.
Lack of this capability makes automated principles or patterns update a recipe for more bloat.