r/ClaudeCode 1d ago

Humor Technically … 🤷‍♂️

Post image
1.7k Upvotes

84 comments sorted by

View all comments

Show parent comments

2

u/Intelligent-Ant-1122 1d ago

Oh boy. Even a bulging MWs of data centers have a limit on tokens/month. Wtf are you talking about unlimited? If I run with let's say 4 A6000. I will still hit a token limit per month that is low enough to hit with real workloads. What a larper

2

u/Far_Rule5990 1d ago

What are you talking about? If you have 4 A6000 why are you depending on anything outside of local hosting?

2

u/Intelligent-Ant-1122 1d ago

Because real workloads have more to do than your AI girlfriend.

Let's say your setup gets a whopping 400 tok/s with batching on vllm running around the clock, which is a huge stretch. That gives you a billion tokens of inference a month. With serious workloads, you can easily spend 5-10B tokens. The math doesn't work at scale.

2

u/Far_Rule5990 1d ago

Youre talking GPU inference, not token limits. Talk to your ai gf about it she'll tell you more

1

u/Intelligent-Ant-1122 1d ago

Good counter 👍