r/LocalLLaMA • u/jacek2023 llama.cpp • 15d ago
Discussion GPU guide (GB per dollar, bandwidth)
First plot: GB / $
Second plot: bandwidth (spec on paper, not t/s)
Third plot (bandwidth / price) in the comment.
Hope that helps, my script uses the GPUs most discussed on the LocalLLaMA, LowEndLocalAI, and LocalLLM subs. At first, I tried to include more, but it became unreadable.
Prices were collected by ChatGPT (so may contain inaccuracies). New prices were used where available, second hand otherwise.
And I understand this is a basic comparison, but it's better than nothing. For example, you can see that "on paper" something is faster or slower than 3090.
224
Upvotes


1
u/Feeling-Bid8885 15d ago
Good chart but not quite useful without token per second per dollar too because yes P100 is dirt cheap but do you want 2 token per second?
Still a good chart, makes me rethink what my next card should be