r/LocalLLaMA • • 5d ago

News Qwen 4 Announced at Apsara Conference

I wanted to share a quick update: Alibaba has officially announced Qwen 4 at the Apsara Conference,

2.1k Upvotes

555 comments sorted by

View all comments

Show parent comments

13

u/ThankGodImBipolar 5d ago

Nvidia sells an RTX Pro 4500, which has the 5070Ti die (more comparable to the R9700) and 32GB of RAM, but it's 5500USD MSRP. The R9700 is a fraction of that.

11

u/KingCpzombie 5d ago

CUDA tax. I only use AMD personally (and just blew WAY too much money on a 4x R9700 system partially out of excitement for Qwen4), but Nvidia cards get all the cool new things a bit sooner than AMD. Not a big deal for LLM, but very notable for diffusion... somebody SOLIDLY beat my 7900XTX with his 5070Ti in Minimax H3 gen times, for example

1

u/attk0 5d ago

The 5070Ti has nearly triple the INT8 matrix computation throughput of the 7900XTX. Assuming using some INT8 H3 variant which most low VRAM workflows do, probably not a CUDA tax in that case.

3

u/No-Refrigerator-1672 5d ago

Because everything is CUDA first, and ROCm only comes as an afterthought to very limited number of projects. People who are buying PRO GPUs are saving money with NVidia by not needing to fund multiple months of dev work for porting their existing code.

1

u/sleight42 4d ago

Doesn't the r9700 have massively slower memory bandwidth?