> and it was running very slowly
... I'm at a loss for words here. It was being served for free. To the entire world.
GPT-5.6 Luna is also served for free to the entire world with a tokens per second rate nearly 10X higher.
> ... I'm at a loss for words here
No need to be so dramatic. I think it's great that they're developing chips, but the whole "RIP nVidia" claim was overly dramatic.
Are you really comparing chatbot to agentic/code work?
Why is Luna not free on OpenRouter? :)
Do you know how much traffic luna was getting vs Ox Alpha?
GPT-5.6 Luna is also served for free to the entire world with a tokens per second rate nearly 10X higher.
> ... I'm at a loss for words here
No need to be so dramatic. I think it's great that they're developing chips, but the whole "RIP nVidia" claim was overly dramatic.
Are you really comparing chatbot to agentic/code work?
Why is Luna not free on OpenRouter? :)
Do you know how much traffic luna was getting vs Ox Alpha?