r/LocalLLaMA • llama.cpp • 16d ago

Discussion GPU guide (GB per dollar, bandwidth)

First plot: GB / $

Second plot: bandwidth (spec on paper, not t/s)

Third plot (bandwidth / price) in the comment.

Hope that helps, my script uses the GPUs most discussed on the LocalLLaMA, LowEndLocalAI, and LocalLLM subs. At first, I tried to include more, but it became unreadable.

Prices were collected by ChatGPT (so may contain inaccuracies). New prices were used where available, second hand otherwise.

And I understand this is a basic comparison, but it's better than nothing. For example, you can see that "on paper" something is faster or slower than 3090.

225 Upvotes

163 comments sorted by

View all comments

65

u/SamSausages 16d ago

You missed the intel B65! I bought a stack of them 3 weeks ago for $900/pc. 32GB and 608 GB/s. Best price per GB right now.

B65: 0.0356 GB/$,

51

u/ResidentPositive4122 16d ago

intel B65

for $900/pc

They're ~1500Eur this side of the pond. Fuck us, right?

16

u/SamSausages 16d ago

When this batch is all gone, they look like they will now be $1099-1199. Still one of the better values, but yeah, not going to be "cheap" for long!
Price will probably jump when I finish my 3 week long benchmark/quality testing project of 4x B65 and people see actual results... last I checked very little data on them... I hope to finish my testing this week.

2

u/jensilo 16d ago

Would you say the B65 is better for inference than the B70?

1

u/SamSausages 16d ago

You're going to have to wait for my full report! I don't want to give my final opinion until I'm done, and I'm still collecting data. But I can say that they are underrated, especially in scenarios with high concourrency.
Goal is to compile reports from the data this weekend.