From opencode go $10/mo plan I get between 60 t/s and 100 token/s even with large contexts of 150k+ tokens.
I wouldn't call 80 t/s slow.
From opencode go $10/mo plan I get between 60 t/s and 100 token/s even with large contexts of 150k+ tokens.
I wouldn't call 80 t/s slow.
You are right, relatively to other llm providers this is not slow. But if you think what is possible when you have 1000t/s a sec you might find it slow.
That's across 64 concurrent streams; you could make more concurrent requests to DeepSeek API no?