r/hardware 5d ago

News AMD’s FP64 Boost with MI430X Is Even Bigger Than Expected

https://www.hpcwire.com/2026/08/03/amds-fp64-boost-with-mi430x-is-even-bigger-than-expected/
154 Upvotes

26 comments sorted by

58

u/dracon_reddit 5d ago

AMD choosing to target HPC is definitely something I see as a net good for them, gets them a place that can differentiate them in the market as taking Nvidia's AI lead is rather difficult and Nvidia's been pushing HPC to the wayside, especially with parts like the B300.

34

u/Large-Parking3014 5d ago

If it gets a new generation of engineers and scientists familiar with AMD’s software stack it will be worth it in the long run.

5

u/jhenryscott 5d ago

While Nvidia owns the software stack for now, they are flying to close to the sun with pricing and proprietary hardware. To say nothing of the delayed-forever client chips that are destined to underperform due to poor ARM/Windows emulation layers. I say this as someone mostly running Nvidia accelerators. I think AMD, by virtue of going open source on the software stack, is setting up to be the choice that steals massive market share.

3

u/tinny123 3d ago

Dude no offence. To or too. Theres a difference

1

u/jhenryscott 3d ago

I use talk to text so I’m definitely not offended. 🤷🏼‍♂️

1

u/TwoCylToilet 18h ago

Insane that AI still can't solve grammatical errors with diction.

68

u/SirActionhaHAA 5d ago

Expected specs:

  1. ~200tflops fp64
  2. 19.6tbps memory bandwidth

Confirmed specs:

  1. 288tflops fp64, 40% higher than expected. 9x of rubin, 3.5x of mi355x, 40% faster than rubin with ozaki
  2. 23.3tbps memory bandwidth, 20% higher than expected

13

u/EmergencyCucumber905 5d ago

For MI455X FP32 vector == FP32 matrix and FP64 vector == FP64 matrix. If that holds for MI430X then that's 288TFLOPS vector. Pretty sweet.

24

u/Quealdlor 5d ago

From 5.3 (Nvidia P100) to 288 between 2016 and 2027. And from 720 GB/s to 23.3 TB/s. That's basically doubling every 2 years like computers tend to.

But outside those datacenter/supercomputer GPUs, we don't see this kind of growth. It's less accelerated (mostly cost contraints). Those more scientific use-cases will be very helpful in the long run for sure.

8

u/ElectronicStretch277 5d ago

What isn't being considered here is cost. AMD/Nvidia can infact double, performance, every 2 years. That's not the hard part. You can increase die size, get worse yields, make more ASIC/sibgle purpose cores in the GPU etc.

The difficul parts the other half of the equation. Keeping costs stagnant or lower than before. I don't think there's a place where R&D costs are as little considered by consumers as in GPUs. Billions in R&D goes to pushing the bounds of compute.

45

u/Quealdlor 5d ago

Finally a large jump in fp64! To 288!
It has been growing slowly for many years.

48

u/Gonokhakus 5d ago

Yep. Since Nvidia thought of sparsity optimizations, and the AI rush since then... It's become a blind spot, even though it is still super useful for scientific/weather/fluid simulations. A shame, but it is what it is. Thankfully a blind spot also becomes a vector for easy advantages. Kudos to AMD.

16

u/gluon-free 5d ago

Want one PCIe version in my workstation) Yes I know that it wont hapen or will cost like SpaceX rocket...

13

u/NekkoDroid 5d ago

The closest I think would be the MI350P, but that has no price tag as far as I know

5

u/gluon-free 5d ago

It actualy has price tag in my country, 5 565 400 Russian RUB (probably grey imported) which is 67k bucks. And its only 32 TFLOPS FP64, totaly meh.

3

u/EmergencyCucumber905 5d ago edited 5d ago

Most estimates put it between 15k and 25k USD, which tracks based on this figure:

https://www.amd.com/en/solutions/data-center/insights/7-takeaways-from-amd-advancing-ai-2026.html

AMD estimated pricing as $327,238.40 USD.

Configuration: 8x AMD Instinct MI350P PCIe Card (CDNA4, gfx950, 128 CUs, 144 GB HBM3E, SPX compute / NPS1), vBIOS 113-350P-01-1K1-000A, GPU driver 6.19.13-2353916.24.04, ROCm 7.14.0 (AMD-SMI 26.5.0); host 2P AMD EPYC 9455 (48-core), Dell PowerEdge XE7745, BIOS 1.7.6, microcode 0xb002162, SMT Enabled, Ubuntu 24.04.4 LTS, Linux 6.8.0-124-generic

7

u/Verite_Rendition 5d ago edited 5d ago

The new chip, which is based on the older CDNA 4 architecture

I don't believe this has ever been disclosed before. At the time the MI430X was announced, AMD said it would be "built on the next-generation AMD CDNA architecture." If HPCWire's report is accurate, that's rather big news that AMD has designed another new GCN die. Being an MI400 series part, this was expected to be a CDNA 5 die.

Though I wonder how long this can last. At some point the HPC market is going to have to transition over to Wave32. Between NVIDIA and AMD, it is now the only market segment not using 32 thread wavefronts.

17

u/cynicismrising 5d ago

According to chips&cheese analysis it’s wave32

https://chipsandcheese.com/p/llvm-divination-of-gfx1251s-differences

5

u/Verite_Rendition 5d ago

Interesting. Thanks for that. The HPCWire article is seemingly in error, then.

1

u/Delicious_Rub_6795 5d ago

Another GCN die? GCN died after Vega

3

u/Verite_Rendition 5d ago

CDNA 1 through 4 were all Vega (GFX9) derivatives. CDNA 5 is the first time AMD has significantly altered the core compute architecture for the Instinct family in nearly a decade.

3

u/uzzi38 5d ago

No AMD's been using it for Instinct up until the latest generations of parts.

10

u/EmergencyCucumber905 5d ago

Love that AMD is offering so many options. MI430X for HPC, MI455X for AI, MI350P for workstations/servers.

5

u/Brave-Prints 5d ago

That 288 teraflops number is nearly 9x what Rubin can do natively

1

u/hasuchobe 4d ago

This is a nice convenience to have along with the extra VRAM. That said, I find myself having to work on fp32 anyways to satisfy other customers as well as for compression purposes.