> Even as someone using AI on the regular I'm starting to hate the "You didn't actually use this exact most expensive model so your point is invalid" argument.
What the author of the article is doing is dismissing a technology so disruptive that it's basically all everyone's talking about in the "tech space" at the moment (I mean look at HN frontpage for the past few months), by trying a relatively mediocre (but still quite good) model for about 10 seconds.
The reality is that frontier models suddenly got very good in the past 3-6 months. It has it's problems, and you need to learn how to use this new tool (as with any tool).
But models can and do generate good code. They also can and do generate absolute garbage (even Fable).
You need a good harness, tools to help the model check it's own output, good context, and a good idea of what you actually want. If you have those 4 things, the chances of generating absolute garbage are pretty slim (but yeah, still there).