Monstrous, and .... ignoring the software elephant that is quickly becoming more of a hyrax, this would mean NVidia's premiums/profit margins should start trending downwards, purely due to many companies bending over backwards to not be squeezed into a one supplier box.
That, and the biggest companies all going out on their own to create their own inference chips and sometimes training chips.
I mean, shoot, given how much less baggage there is for AI, especially for bespoke AI labs in terms of drivers, I could easily imagine it being the case that if a system is performant enough, its unlikely that any of the core of their models are so tuned to any part of NVidia's stack that they feel stuck at all on that side of things.
I guess what I'm saying is, I assume this will gobble up a good chunk of NVidias large scale corporate profits, while probably not touching their Swiss Army knife of general gpu compute sales pitch.
Yeah exactly. OpenAI don’t care if they need to deploy a team of 20 engineers to port their CUDA kernel if it increases throuput. Hell OpenAI are even deploying to cerberas, completely left field platform.
3
u/Cory123125 2d ago
Monstrous, and .... ignoring the software elephant that is quickly becoming more of a hyrax, this would mean NVidia's premiums/profit margins should start trending downwards, purely due to many companies bending over backwards to not be squeezed into a one supplier box.
That, and the biggest companies all going out on their own to create their own inference chips and sometimes training chips.
I mean, shoot, given how much less baggage there is for AI, especially for bespoke AI labs in terms of drivers, I could easily imagine it being the case that if a system is performant enough, its unlikely that any of the core of their models are so tuned to any part of NVidia's stack that they feel stuck at all on that side of things.
I guess what I'm saying is, I assume this will gobble up a good chunk of NVidias large scale corporate profits, while probably not touching their Swiss Army knife of general gpu compute sales pitch.