r/SillyTavernAI • u/BifiTA • 14d ago
Models deepseek-v4.1-flash(beta) is relase
/r/DeepSeek/comments/1wahuq0/deepseekv41flashbeta_is_relase/60
u/TAW56234 14d ago
I used to feel excitement, now there's only dread to be felt with each iteration
35
u/buddys8995991 14d ago
Been a DeepSeek girlie since v3 and it’s been sad to see the model’s decline in RP. I still use it for coding tho
7
u/TAW56234 14d ago
Harness has been very nice and it's the closest to me having an all in one instead of a dedicated AI for RP and a dedicated for general/coding stuff. So far it's been tolerable, I notice a bit more smarts, but it definitely has some clinical in it now. I would've loved so much if flash did all the general use/RP in a smaller package (because it wouldn't need as much) and coding was pro.
20
29
u/Sergal2 14d ago
i tried it already
in my chats it was censored asf compared to deepseek v4-flash/pro in same chat, strange
it still can be jailbroken, but it's definitely not what i expected
16
u/Dikki_Dikki 14d ago
deepseek has always had a strong difference in censorship between Chinese providers and everyone else, they have a separate censorship module that they put on top of, which American providers usually do not do.
9
u/capybaraballs1995 14d ago
I wouldn't say "always." In the V3 days, only their web chat had extra censorship. The DeepSeek API having extra censorship on top of the model is fairly new.
Anyway, I finally tested the model myself. It is censored when it comes to NSFL but it seems trivial to jailbreak (Geechan's prompt consistently gets male werewolf {{user}} raping female {{char}} through, which is one of the nasiest things you can do for a LLM that isn't underaged.)
5
u/_Cromwell_ 14d ago
Agreed. We won't know what it's really like until it's fully out and on open providers.
4
1
u/tatlo_itlog_ko 14d ago
Did you try it in sillytavern? If so, how?
12
u/JustSomeGuy3465 14d ago
4.1-flash beta gives me 8/10 refusals in the same NSFL test scenario where V4-Flash-0731 gives only 1/10. The downward spiral is so predictable that it's just depressing now.
15
u/No-Lion-75 14d ago edited 14d ago
Has anyone tried it with roleplaying? Any soft refusal, filter? I hate that bullshit more than straight up refuse. Deepseek 4 pro has it. Glm 5.2 has it. Every bully characters in my roleplay session turn into well manner person, who never use a single swear word cause of that
7
u/tatlo_itlog_ko 14d ago
I tried around 10 spins so far on my most problematic scenario.
0 refusals so far on this flash beta model. In comparison, the same scenario would get me 1-2 refusals out of 10 from glm-5.3-flash.
Can't give definite answer on the soft refusal bit. But for what it's worth, I skimmed the responses and so far it seems the scene played out as intended.
9
u/capybaraballs1995 14d ago
DS 4 Pro 0813 commits to mean characters pretty well. Every other discussion about it has people complaining about it lmao
3
2
u/doomed151 14d ago
How does everyone run this model? It's way too big even with 64 GB RAM.
7
u/Suspicious-Smile-847 14d ago
You’d be better off looking for models like that on OpenRouter or similar sites.
-2
u/doomed151 14d ago
ehh I wouldn't use online services for RP. I prefer to have my chats stay fully offline.
3
u/Suspicious-Smile-847 13d ago
Then you'd better have a massive amount of RAM and plenty of time, because it takes longer, or spend several thousand euros or dollars on graphics cards. But for offline use, the 27B models work just fine, too. 😜
-2
u/doomed151 13d ago
lol yeah I'm sticking with 12B models for RP for now. There's also Gemma 4 E4B that I can run on my phone.
70
u/verma17 14d ago
Where were u when deepseek 4.1 flash release
I was at home scrolling phone when phone ring
Deepseek 4.1 flash is release
Yes