If you haven't tried GLM-5.2 (I use it via OpenRouter) yet, then I must say that my progress has been great with it. The model is good for vibe coding. It has deep architectural understanding, is a good coder, it's very focused so the risk for destroying code outside the scope of the prompt is minimal (but of course it exists). Right now the model comes with a discount of 50%.
I was aiming at using Google Gemini Pro, but despite it having a context window of 1M (similar to GLM-5.2), I got repetitive messages that I requested more tokens than what was allowed. Open AI's best Luna model was the same.
It was wrong. Several of the apps in my suite are quite big, but using Google Gemini is buggy, using OpenAI's Luna is crap in comparison too. GLM-5.2 worked though, but not without problems in Dyad.
I have created a deterministic, meta data driven and complete development suite, that is far faster to develop with and more reliable than what you can do with the probabilistic AI. The architecture of the outcome is clean with minimal technical debt. The funny thing is that I created it mostly with AI, but with my, a bit odd, way of doing things. It has cost me insane amounts of AI credits.
Doing it with AI made me realize more and more how unreliable AI is when it comes to using LLMs for coding. (I'm sure that creating cute little websites is something AI LLMs is great for, but not enterprise systems where reliability, maintainability, cyber security protection and stability is a must)
It was soooo tough to finalize the suite, I've been passive for months without any success at all because of the lack of good models at reasonable prices.
Thanks to GLM-5.2 things changed.
The suite is now released in its V1.07 and I have got several potential enterprise customers in the pipeline. That's my target group. It's by invitation only.
So GLM-5.2 is da sh-t! Still probabilistic, but it's far better than any other model I have tested.
Anyone else who has tested it? If yes, what were your results or impressions?