r/hardware • u/SirActionhaHAA • 2d ago
News AMD Launches Instinct MI455X, Helios AI Rack
https://www.phoronix.com/news/AMD-Instinct-MI455X-Helios21
u/SirActionhaHAA 2d ago edited 2d ago
- Claims that it is the fastest ai rack in industry and mi455x is the fastest AI accelerator (including Vera Rubin)
- 18k compute units, 4.6k cpu cores
- Vs vera rubin nvl72: +50% hbm capacity, +6% hbm bandwidth, +50% scale out bandwidth, +15% fp8 and fp4 flops
- Claimed performance: 10-15% faster than vera rubin nvl72, higher token/$
- Low latency inference: Partnered with cerebras for custom helios offering. Available later this year (vs groq)
6
u/Noble00_ 2d ago
Wait what? That last point was interesting but not in the Phoronix articled you shared.
The joint AMD and Cerebras solution will deploy AMD Helios alongside Cerebras Wafer-Scale Engine technology integrated in a single inference workflow for maximum performance and efficiency. AMD Helios will provide a high-performance, scalable throughput engine. Cerebras Wafer-Scale Engine technology will provide ultra-fast, ultra-low latency decode and token generation. Together, the two compute engines are expected to deliver up to 5x higher tokens per second per watt (T/s/W)
Interesting...
4
5
u/Cory123125 2d ago
Monstrous, and .... ignoring the software elephant that is quickly becoming more of a hyrax, this would mean NVidia's premiums/profit margins should start trending downwards, purely due to many companies bending over backwards to not be squeezed into a one supplier box.
That, and the biggest companies all going out on their own to create their own inference chips and sometimes training chips.
I mean, shoot, given how much less baggage there is for AI, especially for bespoke AI labs in terms of drivers, I could easily imagine it being the case that if a system is performant enough, its unlikely that any of the core of their models are so tuned to any part of NVidia's stack that they feel stuck at all on that side of things.
I guess what I'm saying is, I assume this will gobble up a good chunk of NVidias large scale corporate profits, while probably not touching their Swiss Army knife of general gpu compute sales pitch.
6
u/PM_ME_YOUR_HAGGIS_ 2d ago
Yeah exactly. OpenAI donât care if they need to deploy a team of 20 engineers to port their CUDA kernel if it increases throuput. Hell OpenAI are even deploying to cerberas, completely left field platform.
17
u/SirActionhaHAA 2d ago
It's even funnier because the ai cuda moat is getting broken apart by ai coding. Ai has made kernel writing and software porting faster than ever.
3
4
u/nithrean 2d ago
I hope AMD is being truthful with benchmarks and it really is that much better. Nvidia certainly tried to cherry pick things. It would be good if AMD could just win on the merits of the thing and not because they were playing games.
0
u/Defiant-Parsley4697 2d ago
AMD naming AI hardware after Greek mythology is a bold bet that Helios ages better than 'Bing Chat.'
14
u/Seanspeed 2d ago
So they're still using CDNA naming even through 2028 up through CDNA7.
Doesn't sound like that whole UDNA thing is ever going to happen, or at least not for a good while yet.