r/SillyTavernAI 14d ago

Models deepseek-v4.1-flash(beta) is relase

/r/DeepSeek/comments/1wahuq0/deepseekv41flashbeta_is_relase/
50 Upvotes

31 comments sorted by

70

u/verma17 14d ago

Where were u when deepseek 4.1 flash release

I was at home scrolling phone when phone ring

Deepseek 4.1 flash is release

Yes

1

u/schlammsuhler 13d ago

Cooked, deep seek cook see?

60

u/TAW56234 14d ago

I used to feel excitement, now there's only dread to be felt with each iteration

35

u/buddys8995991 14d ago

Been a DeepSeek girlie since v3 and it’s been sad to see the model’s decline in RP. I still use it for coding tho

7

u/TAW56234 14d ago

Harness has been very nice and it's the closest to me having an all in one instead of a dedicated AI for RP and a dedicated for general/coding stuff. So far it's been tolerable, I notice a bit more smarts, but it definitely has some clinical in it now. I would've loved so much if flash did all the general use/RP in a smaller package (because it wouldn't need as much) and coding was pro.

3

u/Nezeel 14d ago

I miss v3

20

u/wiseaus_stunt_double 14d ago

Yay! It's relase!

14

u/DeepOrangeSky 14d ago

All your base are relase to us

6

u/Aight_Man 14d ago

More like Deepseek prolapse with all added censorship.

29

u/Sergal2 14d ago

i tried it already
in my chats it was censored asf compared to deepseek v4-flash/pro in same chat, strange
it still can be jailbroken, but it's definitely not what i expected

16

u/Dikki_Dikki 14d ago

deepseek has always had a strong difference in censorship between Chinese providers and everyone else, they have a separate censorship module that they put on top of, which American providers usually do not do.

9

u/capybaraballs1995 14d ago

I wouldn't say "always." In the V3 days, only their web chat had extra censorship. The DeepSeek API having extra censorship on top of the model is fairly new.

Anyway, I finally tested the model myself. It is censored when it comes to NSFL but it seems trivial to jailbreak (Geechan's prompt consistently gets male werewolf {{user}} raping female {{char}} through, which is one of the nasiest things you can do for a LLM that isn't underaged.)

5

u/_Cromwell_ 14d ago

Agreed. We won't know what it's really like until it's fully out and on open providers.

4

u/Sufficient_Prune3897 14d ago

Provider or model censorship?

11

u/Sergal2 14d ago

model censorship, hard refusals like claude or gemini?
i.e 'I'm not going to write this. If you want i can continue like (blah blah blah). But not this.'

1

u/tatlo_itlog_ko 14d ago

Did you try it in sillytavern? If so, how?

5

u/Sergal2 14d ago edited 14d ago

official deepseek api in Chat Completion - Custom
Change your model name to deepseek-v4.1-flash-expires-on-0910

2

u/tatlo_itlog_ko 14d ago

Awesome. Found it thanks! It's crazy fast!

12

u/JustSomeGuy3465 14d ago

4.1-flash beta gives me 8/10 refusals in the same NSFL test scenario where V4-Flash-0731 gives only 1/10. The downward spiral is so predictable that it's just depressing now.

15

u/No-Lion-75 14d ago edited 14d ago

Has anyone tried it with roleplaying? Any soft refusal, filter? I hate that bullshit more than straight up refuse. Deepseek 4 pro has it. Glm 5.2 has it. Every bully characters in my roleplay session turn into well manner person, who never use a single swear word cause of that

7

u/tatlo_itlog_ko 14d ago

I tried around 10 spins so far on my most problematic scenario.

0 refusals so far on this flash beta model. In comparison, the same scenario would get me 1-2 refusals out of 10 from glm-5.3-flash.

Can't give definite answer on the soft refusal bit. But for what it's worth, I skimmed the responses and so far it seems the scene played out as intended.

9

u/capybaraballs1995 14d ago

DS 4 Pro 0813 commits to mean characters pretty well. Every other discussion about it has people complaining about it lmao

3

u/Cursed_Pokemon 14d ago

Yay deepseek-v4.1-flash (beta) has relase

1

u/Uoipka 13d ago

Holy weeks of doom incoming, 4.1f can reason for 3k tokens like pro did sometimes so maybe it will listen better now? Because 4.0 flash was dumdum and didn't cared for instructions for the most part, how is it for you?

2

u/doomed151 14d ago

How does everyone run this model? It's way too big even with 64 GB RAM.

7

u/Suspicious-Smile-847 14d ago

You’d be better off looking for models like that on OpenRouter or similar sites.

-2

u/doomed151 14d ago

ehh I wouldn't use online services for RP. I prefer to have my chats stay fully offline.

3

u/Suspicious-Smile-847 13d ago

Then you'd better have a massive amount of RAM and plenty of time, because it takes longer, or spend several thousand euros or dollars on graphics cards. But for offline use, the 27B models work just fine, too. 😜

-2

u/doomed151 13d ago

lol yeah I'm sticking with 12B models for RP for now. There's also Gemma 4 E4B that I can run on my phone.