It should be very clear that Nvidia is not interested in selling you GPUs to run models. But it’s very interested in selling data centers that you can use to run models in a way that everybody on their end “accepts”.
DGX Spark, RTX Spark, Pro Blackwells, etc. They are absolutely interested in selling local hardware you own. Just at the similar margins to the rest of their main business now.
More the opposite, too many companies buying 5090s by the pallet instead of their Pro or DC lineup.
Just from basic maths, the GeForce line (except 5090) has and continues to be extremely ‘subsidised’ at street prices compared to their DC lineup; in terms of margins, and $ they’d make instead of using the same fab or memory allocation for e.g. a RTX Pro or higher; which still have massive demand.
If you want 32GB and CUDA, 2x 5060ti 16gb and tensor parallelism 2 is pretty easy to set up in any ATX PC case.
They are a greedy mega cap business like everyone else. But if you rationalise it a bit, if they make geforces too compelling for scalers and neoclouds with more VRAM, the margin shortfall would be in the hundreds of billions.
At least they’re not exiting the consumer market altogether like other players, and at least they continue to deliver on software.
It should be very clear that Nvidia is not interested in selling you GPUs to run models. But it’s very interested in selling data centers that you can use to run models in a way that everybody on their end “accepts”.
DGX Spark, RTX Spark, Pro Blackwells, etc. They are absolutely interested in selling local hardware you own. Just at the similar margins to the rest of their main business now.
They just took the 5090 of the market. Too powerful for ordinary people I guess.
More the opposite, too many companies buying 5090s by the pallet instead of their Pro or DC lineup.
Just from basic maths, the GeForce line (except 5090) has and continues to be extremely ‘subsidised’ at street prices compared to their DC lineup; in terms of margins, and $ they’d make instead of using the same fab or memory allocation for e.g. a RTX Pro or higher; which still have massive demand.
If you want 32GB and CUDA, 2x 5060ti 16gb and tensor parallelism 2 is pretty easy to set up in any ATX PC case.
They are a greedy mega cap business like everyone else. But if you rationalise it a bit, if they make geforces too compelling for scalers and neoclouds with more VRAM, the margin shortfall would be in the hundreds of billions.
At least they’re not exiting the consumer market altogether like other players, and at least they continue to deliver on software.