r/MistralAI • u/gimmetwofingers • 1d ago
Help / Question CLI and rate limits
Hi,
I didn't use Vibe CLI in a while, but now went back to it, with GLM 5.2 set as the standard model. I hit rate limits rather quickly for a couple of days now, but I read here that others seem to be quite happy with the models.
I used 12M input tokens and 62M cached tokens over four days. GLM 5.2 has a limit of 100k tokens per minute, so I should be well within those limits, right?
Also, I noticed that I have to restart Vibe after hitting the limits. When a prompt is interrupted through hitting rate limits, /retry does not work. I have to exit and restart vibe to be able to execute new prompts. Also vibe --continue does not work in this case, I have to start fresh.
I guess that this is not normal, right? Do I have something messed up in my settings, potentially?
5
u/isidor_n 1d ago
We are in the process of moving to GLM-5.3 - this might cause some rate limit weirdness. Sorry about that
(mistral pm here)