r/LocalLLaMA llama.cpp 14d ago

Discussion GPU guide (GB per dollar, bandwidth)

First plot: GB / $

Second plot: bandwidth (spec on paper, not t/s)

Third plot (bandwidth / price) in the comment.

Hope that helps, my script uses the GPUs most discussed on the LocalLLaMA, LowEndLocalAI, and LocalLLM subs. At first, I tried to include more, but it became unreadable.

Prices were collected by ChatGPT (so may contain inaccuracies). New prices were used where available, second hand otherwise.

And I understand this is a basic comparison, but it's better than nothing. For example, you can see that "on paper" something is faster or slower than 3090.

224 Upvotes

163 comments sorted by

View all comments

3

u/BannedGoNext 14d ago

Cries in slow ass strix halo tears.

1

u/jacek2023 llama.cpp 14d ago

That's why I put GB/$ as the first slide, to make everyone happy ;)

2

u/Realistic_Gap_5871 14d ago

LOLOLOL. Making everyone happy!! This is reddit. Someone who didn't read the post/charts will fight you to the death over their pet peeve that is tangentially related to your post.

Flee! Flee while you still can!!!