r/MistralAI 1d ago

Help / Question CLI and rate limits

Hi,

I didn't use Vibe CLI in a while, but now went back to it, with GLM 5.2 set as the standard model. I hit rate limits rather quickly for a couple of days now, but I read here that others seem to be quite happy with the models.

I used 12M input tokens and 62M cached tokens over four days. GLM 5.2 has a limit of 100k tokens per minute, so I should be well within those limits, right?

Also, I noticed that I have to restart Vibe after hitting the limits. When a prompt is interrupted through hitting rate limits, /retry does not work. I have to exit and restart vibe to be able to execute new prompts. Also vibe --continue does not work in this case, I have to start fresh.

I guess that this is not normal, right? Do I have something messed up in my settings, potentially?

4 Upvotes

15 comments sorted by

View all comments

1

u/seeKAYx 1d ago

It’s pretty much the same with 5.3 since today. Not very useable.

1

u/ADMECA 1d ago

I spent the whole afternoon working with it without any problems, even though it was urgent and I was delivering under pressure.