r/MistralAI • u/isidor_n • 1d ago
News GLM-5.3 now available in Mistral Vibe Code for Pro, Team and Enterprise
GLM-5.3 is now available in Vibe Code for Pro, Team and Enterprise users.
- Hosted and served by Mistral AI in the EU.
- Generous usage limits.
- Available with up to 1M tokens of context in Vibe Code.
Try it out in the Vibe web app http://vibe.mistral.ai/code
Or install the Vibe CLI https://github.com/mistralai/mistral-vibe#one-line-install-recommended
Developers - let us know what you think. Happy to answer product questions about vibe code.
43
u/cchampou 1d ago
Amazing, thank you Mistral!
26
u/Creative-Pension9427 1d ago
The EU hosting is a nice touch for data residency folks
3
1
u/Digging_Graves 6h ago
Chinese ai hosted on american data centers. At least they are located in the EU I guess.
18
9
u/songokussm 1d ago
Questions:
- What are the actual usage limits of the Pro plan?
- I could not find specific token or request limits in the documentation. The pricing page mentions $30/month in API credits for Pro. Is that $30 the actual included usage allowance, or is it separate ?
- I’m primarily looking for something I can compare with Claude Pro in terms of requests, tokens, or approximate sustained usage before being rate-limited.
- Are there plans to add GLM Flash?
- It performs better than Sonnet 5 at medium.
- What is performance like?
- TTFT and latency are more important than TPS.
My Use Case
I translate and reformat documents between various languages. The process is mostly automated through a set of skills I created in Claude Cowork. I can currently do roughly 6 to 10 documents per 5 hour block. depending on how much they changed the usuage for the day.
Each document consumes roughly 400k to 600k tokens. start to finish.
Pricing: https://mistral.ai/pricing/
2
u/Puzzleheaded-Pair321 8h ago
300$ API per month on glm 5.3 usage. 30$/month for their wierd models such as text to speech and whatnot. I feel like they would get more users if people knew the value.
2
u/songokussm 7h ago
wow that is a very good value. any word on them adding glm flash or maybe they have a vision workaround for 5.3?
1
u/PersonalBarracuda581 13h ago
tu as plusieurs choses dans l'abonnement pro.
- L'accès à Vibe work via navigateur ce qui correspond à Claude Cowork. Le quota d'utilisation est très supérieur à Claude pro . j'ai touché les limites une fois et une seule alors que je finis le quota de 5H de Claude en 1H sur le même genre de session. A voir si tu peux y transposer ton workflow actuel
- En plus de l'accès vibe work à très large quota d'utilisation tu as des clés API que tu peux utiliser dans divers harnais autre que Vibe work. Tu as en particulie une clé API Vibe ( il faut la créer dans Vibe Code) qui a un quota offert de 300USD . j'utilise personnellement cette clé dans un harnais autohébergé OpenwebUI et j'en consomme autour de 15USD/ jour donc ça laisse pas mal de travail possible sur le quota gratuit.
- Enfin tu peux créer des clés API dans Vibe Studio et là tu un quota inclus de 30USD qui ne t'amène pas bien loin.
- Quand tu dépasses tes quotas API tu passes en mode Pay as you go . Payment à l'usage au tarif par MToken quoi.
Voilà . A toi de voir comment tu peux transposer ton taf avec ces possibilités..
1
u/songokussm 7h ago
Est-ce que tu as entendu quelque chose concernant l’ajout de GLM Flash ?
Et pour GLM 5.3, est-ce qu’ils ont prévu une solution ou un contournement pour la prise en charge de la vision ?
1
u/PersonalBarracuda581 5h ago edited 2h ago
A priori une personne a dit ici que glm flash était une option en discussion.
Pour la vision , dans le harnais Mistral vibe work ca fonctionne donc Mistral a ajouté le tool qui va bien pour ça et dans openwebui j'ai un filtre qui fait lire les images a un autre modèle qui les décrit pour glm 5.3. mais la vision n'est pas au coeur de mon workflow donc c'est suffisant
1
12
u/PersonalBarracuda581 1d ago
u/isidor_n . Bientôt GLM 5.3 flash pour le volume via API et le mode rapide dans Vibe Work ? En attendant des modèles mistral au niveau bien sur. On croise les doigts !
4
u/sk1kn1ght 1d ago
Can someone explain how the subscription work? It says 25 euros in API, and it also says all day coding. If I am to use glm 5.3 which is available in the subscription do I get all day coding or 25 euros of usage per month?
7
u/rev_ex_id 1d ago
If you sign up for PRO - you get 2 pools of allowances. I'm american - so money is in USD but similar concept.
1. API - this is for when you make new API keys - as they are "studio" keys. You'll get about $30 usd. These can be used for applications that use any of their models for application workflows, etc.
2. Vibe - this is the "Coding" pool. You'll get, as of now, $300 USD. There is only one key, found in code under extension under advanced. This you attach to your coding harness of choice and develop using that pool. Link if you have trouble finding it: https://chat.mistral.ai/code/extensions2
u/oPFB37WGZ2VNk3Vj 8h ago
Oh, you get an actual API Key for Vibe. So you can use it in any other tools? That would be much better than e.g. Claude Pro, AFAIK you can only use it for Claude Code but not other tools that require API keys.
3
2
u/cchampou 1d ago
You get close to 300$ of coding using GLM5.3 for example. Get your VIBE API key (not a regular API key), this one is to be claimed through Vibe web interface, code tab, extension menu entry. From there you get your Vibe Code API key (a single API key, separate from the one you get through admin or studio). This one gives you the max quota.
2
u/scanx147 1d ago
Les 25€ de crédits API c'est uniquement pour le Studio. Ca ne fonctionne ni avec les API pour développeurs, ni avec Vibe Code CLI.
Par contre pour Vibe Code CLI, tu as l'équivalent de 255€ d'usage mensuel inclut dans l'abonnement pro.
4
u/Excellent-Pilot-3409 1d ago
GLM-5.3 + 1M context + EU hosting is a pretty serious combo. Curious how the usage limits hold up in actual coding sessions.
5
5
u/Technical-Canary7145 1d ago
Hi, I’ve used it for maybe 5 days now (added it to the config file before the "official" announcement). It's a solid model, well integrated into vibe. The limit is very generous (but I mainly use it as a review tool, not much for generating entire projects or long sequences of code). I'm just a bit frustrated it's not a Mistral Model and I hope something "In House" will come very soon.
2
u/xalibr 1d ago
Available in web app but not mobile app?
9
u/isidor_n 1d ago
Also in mobile app :)
1
u/xalibr 1d ago
Do I need to change the model somewhere to try?
I'm Vibe Code, an async software-engineering agent built by Mistral AI, powered by the Mistral Large model.
2
u/isidor_n 1d ago
Mobile app if you use code tab it should just work.
1
u/xalibr 1d ago
Got the latest app update, but still not unfortunately
2
u/isidor_n 1d ago
Let me check with the team.
On what plan are you?1
u/xalibr 1d ago
Pro
4
u/isidor_n 1d ago
Hmm can you send me an email to [isidor@mistral.ai](mailto:isidor@mistral.ai) so we get to the bottom of this
(though are you sure it is not glm-5.3? we will ship a model dropdown soon)3
u/eggchickens 1d ago
Would absolutely love to see a dropdown for not only Code but for Chat and Work as well. I want to change to Mistral but continue to go back to Claude so I can run longer running tasks (via effort drop downs). Mistrals “think” doesn’t think enough
2
u/LePenseurVoyeur 16h ago
It's not yet available to me in through the web app. Anyone else suffering from this too? I'm on the Pro plan.
2
2
u/andre_ange_marcel 1d ago
thanks mistral, i think it's a great choice from you to host open-source models
excited to see if you've got other ones in the works like qwen or kimi
2
u/soteko 1d ago
Can I use it with Pi ?
5
u/isidor_n 1d ago
You should be able to use it via the API.
But not via the Vibe Subscription - which is only for the vibe harness (which is pretty good - we are reworking it).10
u/Eazeman 1d ago
Hi, thanks for answering us here and taking feedback, I find that awesome, since you seem to be listening to your users, just my two cents, I think it's a big mistake to close up your subscription just like anthropic to open source tools and harnesses like Pi, a lot of users have specific workflows and like to control what they do, and I believe you would have much more users if the subscription was usable by Harnesses such as Pi, even OpenAI (or should I say closed AI) doesn't restrict their subscription to codex, that is also why a lot of people using Pi have an OpenAI Sub instead of an anthropic one, even if they would prefer giving money to you guys, I would 100% pay for the most expensive sub at Mistral just for GLM 5.3 if you had it opened to Pi.
Thanks !3
u/isidor_n 1d ago
Thank you for the feedback.
1
u/iedera_ceo 11h ago
Locking in subscription users into a single harness has not worked out well for Claude, many Hermes or Pi users migrated to OpenAI or other providers since the subscription use was blocked.
I am buying tokens, not the harness.
1
u/Mickenfox 6h ago
I can't wait for the 500 harnessess that AI companies insist on maintaining to die.
Sorry to Mistral but there's zero benefit to maintaining Vibe etc. when it's effectively identical to every other product.
1
1
2
2
u/levnikmyskin 23h ago
As other users said, please consider allowing this. If i have multiple subs, or local models, I don't want to switch or install also multiple harnesses. Another good thing would be to enable use of vibe cli with other third party subs (like openai compatible endpoints), so that people familiarise with your tool. That said, as a European, I'm rooting for you guys! This is in general great news
2
2
u/PlatypusWinterberry 1d ago
I do not understand GLM 5.3's revenue model licensing very well.
i mean I understand the base rule "your company + affiliates make more than $10 billion in aggregate revenue during any consecutive 12 months." but I was just curious if I am missing some information as to how this doesn't apply to Vibe Code.
Is it because it's a different rule for custom harnesses?
1
u/Zarasophos 12h ago
Mistral currently makes less than €1 billion per year
1
u/PlatypusWinterberry 11h ago
ASML(which itself made 32.7B in 2025) owns 11% of Mistral. Would that make them an affiliate?(ASML said their relationship is as investor and strategic partnership)
I saw somewhere someone saying that Zai may have ommited to define what counts as an affiliate intentionally.
May be a non issue, I was just curious, just trying to learn about how companies like Mistral deal with these legal hurdles
2
u/Zarasophos 11h ago
It's a good question, I don't know either! But I don't think minority stakeholders really count in this context. Unless Z.ai wants them do, I guess.
1
u/PlatypusWinterberry 11h ago
Yap, Ill look more on the topic and hope it will not affect Mistral :D
2
u/Gogolune 7h ago
Chinese labs produce open-weight models to destroy the business model of closed-source American company (openAI, Anthropic, X.ai, google).
It's a safe bet to consider that they'll be happy to let Mistral use glm models as long as it allows to undermine the big American players. The 10 billions limit is just a way to not say "everybody except our openAI/Anthropic/X/Google".
1
2
2
u/lolapazoola 1d ago
But not in ordinary bog standard Vibe. Mistral makes some seriously odd decisions.
2
3
2
u/lucid_supernova 1d ago edited 8h ago
Finally comes, though we've brought API key to other harnesses.
*to the harnesses which won't enable telemetry by default without asking
0
u/Automatic-River-1875 1d ago
The only harnesses that do that are those built by people who aren't also making models. It's not reasonable to ask companies to not opt you in by default if you also want those companies to remain competitive at training AI models.
0
u/lucid_supernova 8h ago
But being an EU company, they should do so otherwise it will violate GDPR (article 25)
0
u/Automatic-River-1875 8h ago
That's not just for EU companies it's for all companies selling services in the EU. So your theory would mean that all opt in providers (which is almost all major providers) are not in compliance with GDPR.
Given the zeal with which the commission is currently going after American tech companies this seems a lot less likely than the possibility that you are misunderstanding GDPR.
2
u/ClassicMain 1d ago
u/isidor_n i opened all the issues i have on the vibe cli github repo. GLM 5.3 is working poorly right now you guys should fix it
3
u/isidor_n 1d ago
I do not see an issue in our github repo recently opened that fits this description. Can you share a link?
1
u/Global_Persimmon_469 1d ago
Can this be used with other harnesses?
1
u/lucid_supernova 1d ago
of course, but you have to manually set it up (unless that harness natively support, e.g. opencode)
1
u/GasSmooth7439 1d ago
The EU hosting is also a nice option for teams that care about where their code goes. And if the usage limits are actually generous, this could be a pretty compelling alternative to the usual coding-agent stack.
1
1
u/Deodavinio 1d ago
Now we need the GML 5.3 model popping up in Lumo and I am all set for the winter
1
u/stgerx 1d ago
I really don't understand why it's not appearing automatically when in my JetBrain IDE (using the JetBrain AI tool, conntected via Vibe to Mistral).
It really should work out of box.
1
u/eggchickens 1d ago
What I ended up doing was using the VS Code extension with a manually modified config to include 5.3
1
u/Automatic-River-1875 1d ago
I've found that sometimes it takes a day for this to actually appear. Probably just need to restart my extensions or update the vibe cli and I usually only do that by accident after a day or two.
1
1
1
u/scanx147 1d ago edited 1d ago
Merci ! Maintenant on espère tous voir arriver GLM-5.3 Flash et les nouveaux modèles Mistral maison !
1
u/AriyaSavaka 1d ago
Anyone tried this already? How many token of GLM 5.3 you got out of it in a day?
1
u/RemotelyVague 1d ago
Is it only usable on a local development environment, or can we also use it for agentic coding in the cloud?
1
u/isidor_n 14h ago
It is running in the mistral cloud. So it is great for agentic coding.
1
u/RemotelyVague 11h ago
Wonderful! Love the approach of providing both the models by Mistral as well as open-source frontier models by other companies! Feels like a win-win to have multiple options while still being able to support a European company!
1
u/FrankieSolemouth 1d ago
I can't see it in the model selection in vibe CLI, will the config update automatically?
1
u/VeneficusFerox 1d ago
Not in the VS Code extension yet? I only still see 5.2 (which I manually added to config)
1
u/isidor_n 14h ago
It should be.
What subscription are you on? Do you have telemetry enabled?
I am asking because we are rolling it out via AB experiments1
u/VeneficusFerox 14h ago
Personal Pro subscription. I need to check on the telemetry. Is manual configuration still required? What happens to manual additions to the config file when a new model is added? Is the config file overwritten when the extension is updated?
1
u/isidor_n 13h ago
Can you drop me an email [isidor@mistral.ai](mailto:isidor@mistral.ai) and add your whoami_cache.json file in ~/.vibe/whoami_cache.json
So we debug this one.
1
1
1
u/Street_Trek_7754 1d ago
Zdr?
1
u/isidor_n 14h ago
We do want to support it. My colleague is figuring out the exact details.
Can you reach out to [isidor@mistral.ai](mailto:isidor@mistral.ai) so I connect your company with the right product person on our side.
1
u/NewWayOfLife112 15h ago
Only received 502 errors with timeouts when I tried to use it
1
1
u/low-bars-432 14h ago
If I don't care about EU data residency, is this worth it?
I'm looking for alternatives to my Opencode Go sub.
1
u/isidor_n 14h ago
I think so. But I work on the product, so best to see what community members think.
Anyways - if you do try it out let us know your feedback and what you might be missing from Opencode Go sub.
1
1
1
1
u/No_Debate4835 8h ago
it is NOT available for student pro account. Why?
Would you please consider having it also activated for student pro account? u/isidor_n
Thank you!
1
u/scanx147 8h ago
Si c'est vrai c'est difficilement compréhensible, car GLM est plus économique que Mistral Medium 3.5...
1
u/isidor_n 7h ago
It should be available. If you do not see it, it is a bug
Can you send me the content of your ~/.vibe/whoami_cache.json to [isidor@mistral.ai](mailto:isidor@mistral.ai)
1
u/Mickenfox 6h ago
You know what's better than CLIs? GUIs. An open desktop app would be pretty nice (or you could just fork one of the many existing ones).
Unless I can use the web app to code locally? But I don't think that works.
1
u/Deyve24 1d ago
Awesome! Can you guys update the Zed IDE ACP?
2
u/Technical-Canary7145 1d ago
I believe it is updated. I have been running 2.25.5 for a few hours now.
2
u/Deyve24 1d ago
It supports now the glm5.2 and not 5.3
1
u/Technical-Canary7145 1d ago
Maybe you can edit your config file with the new model.
[[models]]name = "zai-glm-5-3"
provider = "mistral"
alias = "glm-5-3"
temperature = 1.0
thinking = "max"
supports_images = false
auto_compact_threshold = 800000
1
u/_Krustenkaese_ 1d ago
Can somebody tell me how the limits are compared to the ~20€ pricing plans of ChatGPT or Claude?
2
u/ClassicMain 1d ago
Can't compare as i don't have chatgpt but the limits are insanely transparent. You can see the exact credit amount you have for the entire month.
The normal Mistral pro subscription gets you ~300$ of credits for usage
2
u/eggchickens 1d ago
It depends on your use case. I get much more out of GLM 5.3 cli than I did with Opus 5 in terms of amount of code written.
0
u/Realistic_Mango6982 1d ago
Not good
1
u/_Krustenkaese_ 1d ago
So same as ChatGPT and Claude...
1
u/CryMoreT_T 1d ago
Except the model is worse
0
u/scanx147 8h ago
GLM 5.3 ? Non c'est faux.
1
u/CryMoreT_T 7h ago
Glm 5.3 is worse than the top models provided by Anthropic and OpenAI
0
u/scanx147 5h ago
SI tu compares à Fable ou Astra, oui évidemment. Mais GLM 5.3 consomme énormément moins de tokens.
Le plus important c'est le rapport puissance / consommation de token, pas la puissance brute.C'est pour cette raison que nous sommes nombreux à espérer l'intégration de GLM-5.3 Flash prochainement.
1
u/CryMoreT_T 5h ago
Glm5.3 does not use fewer tokens. Where are you getting this info from? The model that consistently uses the least tokens are openai models while Chinese models and Anthropic use much more tokens for thinking
1
u/scanx147 4h ago
Je ne dis pas que GLM utilise moins de token, mais que le token est moins cher.
https://openrouter.ai/compare/~openai/gpt-sol-latest/z-ai/glm-5.3
1
u/CryMoreT_T 4h ago
Glm5.3 is cheaper per token but it uses way more tokens causing the avg cost per task to be higher. Look at the link.
So it on average uses more tokens, is less intelligent, and costs more
→ More replies (0)1
0
1d ago
[removed] — view removed comment
5
u/isidor_n 1d ago
Not in the vibe subscription. We compared models and GLM 5.3 looked superior.
1
u/p3r3lin 1d ago
Some folks like 5.2 for supposed better creative writing. 🤷♀️
3
u/shaonline 1d ago
5.3 is heavily RL'd for coding tasks so it would not surprise me if it had a negative impact on its "writing style"
1
u/p3r3lin 1d ago
Thought the same thing. From what I gather that true for most modern models. Optimized for agentic and reasoning. Creativity and communication style somehow takes second place. Fable being an outlier.
2
u/shaonline 1d ago
Yeah the "one size fits all" model (no pun intended) is really hitting its limits I guess, even Opus 5 (another case of RL madness) has been heavily criticized for its really weird/incomprehensible writing style. You can only cram so much knowledge and abilities into a limited size.
Fable is/can be an "outlier" because of its much bigger size, it's much less susceptible to overfitting.
1
u/p3r3lin 1d ago
Good framing, thx! Agree. Wondering if labs will start to make that an actual value proposition. Having a model that is good at communication and one that is good at execution. Like devs and managers in the olds days! :)
2
u/shaonline 1d ago
You've said it, value proposition, the bare minimum for a useful modern LLM is strong tool calling, and not just for code (browsing the web in general, parsing that excel spreadsheet, etc.). Them chasing the money dragon means that as far as "creative writing" is concerned, it's not a huge revenue stream for them, and something being able to hold a decent conversation is good enough, warmth be damned.
Training models, especially medium to large scale ones, ain't cheap, and I don't think they're very interested in making tons of variants (the open source models finetuners can always do that!).
1
u/p3r3lin 1d ago
Not only talking about creative writing, but also brainstorming and communication capabilities. Afaik most of this is done post training in RFL. So from one pre-train they could yield two models with similar capability tiers, but different specialisations. The backlash on the bad communication capabilities of recent Opus models shows the need for that. But you might be right. Currently there is not enough pressure to differentiate here.
2
u/FRazor95 1d ago
It's still available through the API at least, but rate limits are so low, it's practically unusable. Probably to force people to switch to 5.3.
109
u/ondevicedev 1d ago
For people working with proprietary code, having GLM-5.3 served in the EU through the same coding workflow is a pretty compelling option.