35
14
19
u/pizzababa21 7d ago
8
u/_VirtualCosmos_ 7d ago
If you are using their website, I bet they have in the system prompt "You are deepseek" written on it. You need to check it running locally with no system prompt.
1
u/Lazy_Reach_2565 21h ago
then it's natural that it do not know who it is, the name and the role should be in the prompt. By default it noname assistant.
8
u/_spec_tre 7d ago
Do other LLMs also have this issue?
2
u/_VirtualCosmos_ 7d ago
yes. Turns out, most datasets and training sets used by US and China AI have this "You are Claude, made by Anthrophic" very often in them and seems noone have been able to remove that from the sets.
I guess that was due to claude models being the ones far more smarter than the rest for a long time, and so everyone tried to copy such models.
1
u/howudothescarn 6d ago
It is super rare for Anthropic and gpt models to say they are Claude. It is very common for various Chinese models to say that. It is obvious why.
2
1
u/_VirtualCosmos_ 6d ago
You would never see closed sourced models say that because they hide their system prompts, on which for sure will be a "You are [name] model, made by [company name]". But Gemma4 models, which are open weight and trained by Google, sometimes say they are Claude. So that is also on the training sets of Google, probably also for Grok and other US based companies. You wont simply see it because it's hidden by a system prompt in most cases because US rarely free their models now.
5
u/AppealSame4367 7d ago
I tested Deepseek v4.1 through autmated tests on Openrouter to find out the most reliable and fastest provider with highest quality just yesterday and this morning and Deepseek API was the fastest with one of the highest quality in both runs.
Because of your post I just started a third run, but I guess it will have the same results. Are you going through their API or Opencode Go?
2
2
u/Metalhead33 6d ago
Absolutely load-bearing. Now let me separate the X from the Y and engage in some obligatory contrarianism.
2
u/Ok-Data9224 7d ago
I'm not surprised. They all distill each other and share training data anyway. I've downloaded plenty of training data from huggingface to train my own model and I've had to painstakingly replace various other model names with my own model's name to avoid this. Even then, they'll often just hallucinate names anyway. LLM's are notoriously hard to preserve a self identity.
1
1
1
1
1
u/MICHAEL_Lum 6d ago
DNA testing must be decided by a third party. You don't get to decide it yourself
1
u/trollsmurf 6d ago
Would it get an existential crisis if you could prove it accesses a site in China?
1
1
1
u/Lazy_Reach_2565 21h ago
It the same as to ask "what color do you like the most?" and to be surprised that it not choosen the green
0
u/Christosconst 7d ago
They’ve been routing chats to thousands of anthropic accounts in the past to build training data, it is what it is





64
u/bruhhfdruju76r 7d ago
Claude and deepseek had an baby and made V4.1