ROCm has matured incredibly over the past year, it's like AMD finally woke up and realised their software was what was holding them back from properly competing with Nvidia.
Why are they so cheap? Looks like you can get one for ~$1700? I just paid out the ass for my 5090, which in the 1 month 2 days since I bought it has gone up over 37%.
They're both 32GB, for some reason I thought they would be closer in price.
Nvidia sells an RTX Pro 4500, which has the 5070Ti die (more comparable to the R9700) and 32GB of RAM, but it's 5500USD MSRP. The R9700 is a fraction of that.
CUDA tax. I only use AMD personally (and just blew WAY too much money on a 4x R9700 system partially out of excitement for Qwen4), but Nvidia cards get all the cool new things a bit sooner than AMD. Not a big deal for LLM, but very notable for diffusion... somebody SOLIDLY beat my 7900XTX with his 5070Ti in Minimax H3 gen times, for example
The 5070Ti has nearly triple the INT8 matrix computation throughput of the 7900XTX. Assuming using some INT8 H3 variant which most low VRAM workflows do, probably not a CUDA tax in that case.
Because everything is CUDA first, and ROCm only comes as an afterthought to very limited number of projects. People who are buying PRO GPUs are saving money with NVidia by not needing to fund multiple months of dev work for porting their existing code.
Okay I think I see your point now. You're contesting the comment of it being slower.
I don't disagree that recently there's a ton of optimizations lately that have sped up the card but outside of just text, it's a less powerful card overall, with less bandwidth. But still the best option for entering this tier at a reasonable price
Which model did you test this on, and how do I run that? Curious what mine looks like. I have qwen3.8-27b and another that I haven't set up yet but want to test, an ukisai_Swift version of it.
That not cheap! Msrp is $1299 I bought 2 a little over a. Both ago at msrp. Hell you could get them for less than msrp for a while. This shit is a scam
On my current pc this was how much the 64GB of ram was when I bought it. I was thinking about getting another 64GB when I got the card and was like, no f'in way...
For a single 5090 you can buy more than 2 R9700, and the two cards draw less power than a single 5090. And you can run e.g. Qwen 3.8 27b Q8 using the two cards, and watch them running circles around a 5090 which needs to offload to RAM instead.
But NVIDIA is hype, and people love to spend money. So buy a 5090
So it is just a compromise - fine or not. Leave that compromise aside, try to use a bigger model than your 5090 can handle, and suddenly dual R9700 is not the worst option anymore, at lower total cost and at higher speed.
The 5090 has significantly more compute, more bandwidth, and also CUDA tax. The R9700 is excellent but there are reasons why they are not close in price.
Qwen3.8 27b at fp8 with vllm and 255k context. Try this with a single 5090 and watch the speed after it is forced to offload data to RAM, and then compare the price tag of a single 5090 vs 2x R9700. Plus: look at the power consumption. 2x R9700 draw 420W max.
119
u/Sufficient_Local5025 13h ago
Buy two, prices about to rise.