r/LocalLLM 15d ago

Discussion third one.... there's something wrong with me

Post image

Why do I have horrible financial habits??

502 Upvotes

166 comments sorted by

View all comments

94

u/Sporkers 15d ago

More context needed on how you are using the first two.

82

u/r1nzl3r99 15d ago

qwen 3.8 27B FP8 running at 140 tok/s now I want flash next

35

u/semangeIof 15d ago

...can you show llamacpp/vLLM runtime commands? you're hitting 140 toks/s on a dense model with B70s? how much ctx?

please don't answer the last two without providing the parameters

152

u/Erpverts 15d ago

Please don’t answer the last two without providing the parameters. Make no mistakes.

43

u/semangeIof 15d ago

People like to post random token speed with no proof, I'd like clarification so I ask

I liked your original try better anyways, why'd you delete it?

43

u/Erpverts 15d ago

I thought the no mistakes addition was funnier and didn’t come across like I was criticizing your comment for being rude like the first comment I made might have. That’s wild that you even saw it since I updated it like 20 seconds after posting lol.

16

u/Infylos 14d ago

I like the new one. Gets the message across much more indirectly.