You want to solve a family of problems using some tool. You measure the relevant solutions performance metrics.
Now, you or someone else vibe-coded a new tool. You created new solutions and measure again.
You have just quantified the result of using the new tool.
How do you measure them if you don't understand what you're doing? A shitty benchmark or small test suite is not how solid software gets made.