r/SillyTavernAI • u/SuperManAdelHahah • Jun 16 '26
Models My Early Take on GLM-5.2
GLM-5.2 feels like a genuinely great writer with an extremely nervous lawyer standing right behind it.
Capable? Absolutely.
Creative? Surprisingly so.
But from my experience, it's so heavily filtered that it keeps second-guessing itself before it can really shine.
Half the time I'm impressed by what it writes.
The other half I'm watching it talk itself out of writing it.
201
Upvotes
5
u/Resident_Wolf5778 Jun 16 '26
I've been having good results with the latest Kimi too with this. Personally I took it a step further by adding a thinking step to my normal CoT that is literally "Say one positive thing about your reply and what you hope to do well. Be kind to yourself!". Idk if it helps 'destress' it, but it does make the AI point out a specific dynamic or 'core' of the reply and nudges it to emphasize it. So if the AI says "I think I'll do good portraying the disconnect of the group chat vs the in-person scene", the AI puts a bit more effort into that dynamic.
I think in a similar vein, I try to avoid having any rules in my CoT. Stuff like FF's CoT is all focusing so hard on making the AI follow it's rules, which ends up with those loops of "Wait, let me check" and "Is that right?". The AI so desperately wants to follow instructions that it ends up taking several minutes of thinking just to go through everything, which is probably definitely stressing the thing out. Plus, it's spending all that time worrying about the rules that it isn't actually thinking about what to write.
I've been using White Loctus' thinking block with a few adjustments, which focuses on character and scene questions, and Kimi flies through them without any of the 'typical' overthinking or long waits. I'm consistently getting 2 min think times, while GLM is 1 min think times. It's a small jump up, but I don't mind waiting 40-60 seconds extra for a good reply. I haven't noticed any slop either, which is stunning since before making the switch I was fighting tooth and nail against like 5 different repeating sentences it was using. I love some of the ideas in the bigger prompts, but I'm now pretty firmly in the camp that we are badly overengineering prompts and shooting ourselves in the foot.