I don't see how this could work, right now, since every failing I've had was just generic stupidity of the AI, which is brilliant one minute, and a complete idiot the next.
I think correction actions still require the ability to execute them, which (in all the cases I've had) would require more capable models.
Long term, I think you're probably correct.