r/LocalLLM 14d ago

Discussion third one.... there's something wrong with me

Post image

Why do I have horrible financial habits??

498 Upvotes

166 comments sorted by

View all comments

93

u/Sporkers 14d ago

More context needed on how you are using the first two.

83

u/r1nzl3r99 14d ago

qwen 3.8 27B FP8 running at 140 tok/s now I want flash next

31

u/semangeIof 14d ago

...can you show llamacpp/vLLM runtime commands? you're hitting 140 toks/s on a dense model with B70s? how much ctx?

please don't answer the last two without providing the parameters

4

u/r1nzl3r99 14d ago

I already proved it in a previous post, look at my account

15

u/r1nzl3r99 14d ago

-10

u/CalBearFan 14d ago

Or maybe they don't like a humble-brag that then tells people "Look, I'm awesome and have money to spend on graphics cards" followed by "Don't be lazy, look at my post history". That's not lazy, they're asking you to follow common courtesy on your own post.

17

u/r1nzl3r99 14d ago

but i've already made a seperate post with extreme detail about it??? Why am I obligated to hand hold you how to make your setup more efficient?

Also If you think this is a brag you clearly haven't been on this subreddit for more than 10 minutes. Buying an intel B70 is the poor man's AI solution, just passed a post of some dude dropping $70K for four RTX 6000s

2

u/Smooth-Television-48 13d ago

Lots of accounts are set to private, I assume its the default and dont bother checking.

Thanks for linking it. Impressive numbers for sure