True, but it would have to be more than massive (order(s) of magnitude) to offset that gap.

On some benchmarks models like Qwen 3.8 Max which cost < $6/m out cost more than Astra 6 to run at $50/m out. That’s a huge price gap and yet Astra would be cheaper if your work looks like the benchmark.

We notice with frontier models like Astra and Fable that one might use a lot less tokens than the other to complete the task thereby being the better deal in spite of the far higher token cost.

What is Astra $/task?

Even if OpenAI end up using 1 token for per task, if the token costs 1M$ , some people will find it expensive.