> You can ask a model for output directly and stop
That’s still several orders of magnitudes too slow to fit fast vs slow. Think of 30ms vs 3-4 seconds to get an idea of what we’re talking about here
That’s a function of the amount of processing power involved not the underlying architecture of decision making.
That’s a function of the amount of processing power involved not the underlying architecture of decision making.