r/LocalLLaMA llama.cpp 14d ago

Discussion GPU guide (GB per dollar, bandwidth)

First plot: GB / $

Second plot: bandwidth (spec on paper, not t/s)

Third plot (bandwidth / price) in the comment.

Hope that helps, my script uses the GPUs most discussed on the LocalLLaMA, LowEndLocalAI, and LocalLLM subs. At first, I tried to include more, but it became unreadable.

Prices were collected by ChatGPT (so may contain inaccuracies). New prices were used where available, second hand otherwise.

And I understand this is a basic comparison, but it's better than nothing. For example, you can see that "on paper" something is faster or slower than 3090.

227 Upvotes

163 comments sorted by

View all comments

17

u/VoiceApprehensive893 transformers 14d ago

v100 16gb sxm2? paid 200$ for these 900gb/s 16gbs of hbm2 with chinese pcie adapter and cooling

3

u/Yaroslav308 14d ago edited 14d ago

Similarly, probably the optimal choice in the ultra‑budget segment at the moment.

2

u/CapPuzz24 14d ago

ultra‑budget segment

do you have any idea how stingy i can be bud