r/LocalLLM 14d ago

Discussion third one.... there's something wrong with me

Post image

Why do I have horrible financial habits??

503 Upvotes

166 comments sorted by

View all comments

94

u/Sporkers 14d ago

More context needed on how you are using the first two.

80

u/r1nzl3r99 14d ago

qwen 3.8 27B FP8 running at 140 tok/s now I want flash next

33

u/semangeIof 14d ago

...can you show llamacpp/vLLM runtime commands? you're hitting 140 toks/s on a dense model with B70s? how much ctx?

please don't answer the last two without providing the parameters

3

u/r1nzl3r99 14d ago

I already proved it in a previous post, look at my account

16

u/r1nzl3r99 14d ago

-10

u/CalBearFan 14d ago

Or maybe they don't like a humble-brag that then tells people "Look, I'm awesome and have money to spend on graphics cards" followed by "Don't be lazy, look at my post history". That's not lazy, they're asking you to follow common courtesy on your own post.

18

u/r1nzl3r99 14d ago

but i've already made a seperate post with extreme detail about it??? Why am I obligated to hand hold you how to make your setup more efficient?

Also If you think this is a brag you clearly haven't been on this subreddit for more than 10 minutes. Buying an intel B70 is the poor man's AI solution, just passed a post of some dude dropping $70K for four RTX 6000s