You are right, relatively to other llm providers this is not slow. But if you think what is possible when you have 1000t/s a sec you might find it slow.
That's across 64 concurrent streams; you could make more concurrent requests to DeepSeek API no?
That's across 64 concurrent streams; you could make more concurrent requests to DeepSeek API no?