You can ask a model for output directly and stop, or you can recursively ask it to keep refining the output.
That seems to fit the fast vs slow model of human thought reasonably well.
> You can ask a model for output directly and stop
That’s still several orders of magnitudes too slow to fit fast vs slow. Think of 30ms vs 3-4 seconds to get an idea of what we’re talking about here
Not quite. The better analogy for you is the "thinking" setting on your model.