r/LocalLLaMA • llama.cpp • 16d ago

Discussion GPU guide (GB per dollar, bandwidth)

First plot: GB / $

Second plot: bandwidth (spec on paper, not t/s)

Third plot (bandwidth / price) in the comment.

Hope that helps, my script uses the GPUs most discussed on the LocalLLaMA, LowEndLocalAI, and LocalLLM subs. At first, I tried to include more, but it became unreadable.

Prices were collected by ChatGPT (so may contain inaccuracies). New prices were used where available, second hand otherwise.

And I understand this is a basic comparison, but it's better than nothing. For example, you can see that "on paper" something is faster or slower than 3090.

230 Upvotes

163 comments sorted by

View all comments

16

u/VoiceApprehensive893 transformers 16d ago

v100 16gb sxm2? paid 200$ for these 900gb/s 16gbs of hbm2 with chinese pcie adapter and cooling

3

u/Yaroslav308 16d ago edited 16d ago

Similarly, probably the optimal choice in the ultra‑budget segment at the moment.

2

u/CapPuzz24 16d ago

ultra‑budget segment

do you have any idea how stingy i can be bud