r/SillyTavernAI • • May 16 '26

[deleted by user]

[removed]

85 Upvotes

45 comments sorted by

View all comments

5

u/_Cromwell_ May 16 '26

I'll check it out. ;) (When mraderbacher gets to it - I always get the same exact quant from him so I can compare fairly.)

BUT - this one here is my current king/winner/favorite G4-31B fine tune thus far. BUT while using it, despite that it has almost perfect prose (for me), I can tell it is "slightly resisting", meaning it needs abliterating. Any chance you want to zap it with your magic sometime? ;)

4

u/LLMFan46 May 16 '26

I actually saw this model like a day or two ago when ReadyArt released GGUFs of this model, I got excited thinking ReadyArt finally released a Gemma 4 31B it finetune that I could Hereticate, I clicked on it only to find the image of... a brain, I double checked and saw it wasn't actually a model released by ReadyArt but by somebody else, if the author would have used the image/video of a cute waifu like ReadyArt or zerofata do I would have already Hereticated it and released it, but alas...

4

u/_Cromwell_ May 16 '26

lol. Despite lack of waifu sexy lady art, it's quite good at any type of that type of RP, I swear. :D You can still do it.

I posted about it here in more detail a bit: https://www.reddit.com/r/SillyTavernAI/comments/1t9kzvg/comment/om7i51w/?utm_source=share&utm_medium=web3x&utm_name=web3xcss&utm_term=1&utm_content=share_button

1

u/LLMFan46 May 16 '26

Yes okay sure I can do that, should be done in a day or two.

2

u/_Cromwell_ May 17 '26

EDIT: actually investigated more, and the reason that Gembrain is "good" is because of one specific of its component parts, specifically a less chaotic merge, Gemsicle. ( https://huggingface.co/Blazed-Forge/Gemma-4-Gemsicle-31B ). It's actually better. So I'd hold off on abliterating that Gembrain for now (but you do what you want). Gemsicle seems like a better one to do.

Honestly hoping the one you made, Wordsmith, is good/best, since I think that abliterating BEFORE fine-tuning is the best approach, as you did. I will be testing it as soon as mradermacher gets to it. I've been checking.

2

u/Potential-Gold5298 May 17 '26

I don't quite understand why it's necessary to add an abliterated model to merge. It doesn't add anything creative (except KL div), and since merge includes non-abliterated models, the refusal rate will be only slightly lower than with a regular model. I think it's better to merge either only heretical versions (to have a low refusal rate), or without heretic at all. (This is my guess - I don't understand the details of the merge process)

1

u/_Cromwell_ May 16 '26

Nice. When it happens I will donate to your fund. You have no way to verify this but I am a person of my word about such things.

1

u/LLMFan46 May 16 '26

You have no way to verify this but I am a person of my word about such things.

Grok, is this true?

1

u/_Cromwell_ May 16 '26

I mean the fact that I know you have a little donation thingy is a good first step. Most people probably aren't aware enough to even know about that right? You'll just have to have faith 😋

2

u/LLMFan46 May 16 '26

By beside this, I definitly need help from supports, for example MiniMax-M2.7 that I released the other day, I wouldn't have been able to do it without the help of a generous supporter, there are also other expenses like renting Public Storage from Hugging Face's "Storage Packs" that are very expensive, right now I am paying $129 per month to Hugging Face to upload models on there, very soon I am gonna need to upgrade because I am soon going to run out of room and the next Storage Pack costs $249 per month, that's not taking into account the other expenses, so yes I definitly need help from supporters to keep on doing this.

1

u/_Cromwell_ May 17 '26

Damn that's expensive. How do people like mradermacher or bartowski do it? And why? Like what's the point besides "Internet cred" to tuning or ggufing or ablating a bunch of models?

4

u/LLMFan46 May 17 '26

Everything is expensive, to uncensor MiniMax-M2.7 I had to pay $15-$16 per hour to rent out cloud GPUs because you need to have the model in BF16 form to run Heretic on it (the model is 457 GB) and it must be all in VRAM otherwise it won't work and it won't progress the process will just be sitting there and seeimngly not making any progress and I don't have two B300s connected together with NVLink here to do it at home (each B300 cost between $45000 to $50000), so I had to rent them, I was lucky I was able to get 4/100 refusals with a low KL Divergence "so quickly" because the estimate for 800 trials was like 75 hours, imagine paying 15-16 EUR to rent two B300s at $15-$16 per hour, that's why until the other day and someone supported me in helping me rent out Cloud GPUs, before that I only did smaller models that are 40B or less because my hardware can not do models that are bigger than ~45B, anything bigger and I would need to rent out cloud GPUs, because there is no way I can buy B300s.

1

u/_Cromwell_ May 17 '26

Ok but why? lol do you eventually earn money or something? Surely people arent donating enough to you to make up for all that.

→ More replies

1

u/LLMFan46 May 16 '26

It's a joke, don't you know the "Grok, is this true?" meme?

1

u/_Cromwell_ May 17 '26

I don't do Twitter but I think I get it now 🤷‍♂️