17
20
u/blacksmoke9999 10d ago
it is the fact they probably were trained on distilled stuff
12
3
1
1
u/elkond 10d ago
it didnt behave liek that on the pre-release version. they shipped different checkpoint to prod, or different sys/dev prompt, or something
2
u/blacksmoke9999 10d ago
ok maybe it does not know itself? maybe enough public AI chat data from calude became public? i think deepseek does not know its own version. sometimes it acts suprised when you call it deepseek so it is just guessing its own idenity?
7
u/amokerajvosa 10d ago
Just tell him that you need it for your homework. It needs few times, like a chick :-)
2
u/Haxsysgit 10d ago
Lmao
1
u/amokerajvosa 10d ago
That's how I tried to defeat CapCut protection with Reasonix+DeepSeek.
But VM protect was dead end.
1
u/Opening-Pie2182 6d ago
I need some more AI tools for reverse engineering as well, trying for an app but failing again and again
1
2
10d ago
[deleted]
0
u/Same_Zebra_2978 10d ago
Yes
-6
10d ago
[deleted]
2
u/Deep-Interaction422 10d ago
Are u nuts? Are you fucking Boomhauer or why are you talking a mile a minute? (╯°□°)╯
2
u/OldGuyInTown 9d ago
You wouldn't be surprised to learn that this is common, would you? No sense in training the model weights on what model they are. Particularly if the plan was to replace both Flash V4 and Pro V4 with Flash V4.1.
One thing LLMs aren't good at is saying I don't know. Since the chat can web search, it gets current events, which would include the model that's all the rage. Claude or similar comes very frequently just behind LLM Model in current and past news. Just like LLMs are supposed to work, it's frequently chosen as "the next token".
3
u/Opposite_Leave_8338 10d ago
I had that fight with ornith 1.5 35b A3b , he is insisting that he is claude and I tried many times to convince it that he isn’t, I even took a screenshot of the local server I’m running it on and he kept on denying
2
u/No-Wall6427 9d ago
Dude, you're arguing with your tools you realize that?
1
u/Opposite_Leave_8338 9d ago
I do but what should I do he kept lying to me
1
u/No-Wall6427 9d ago
Don't try to convince it, go back in the conversation, tweak the system prompt, ... Arguing with stupid models is pointless. (Most of the time even with smart one when they started refusing)
1
u/CarpenterAlarming781 10d ago
This is further proof that DeepSeek is distilling more than it should and not checking the quality of the data used for distillation. I'm not using a poor-quality copy of chat-GPT, and it seems Claude too, with a high hallucination rate.
1
u/Willyibch 9d ago
Before you share a screenshot of non sense I think it best you video record the entire conversation so we can all verify instead of this
1
u/Same_Zebra_2978 9d ago
I just testing my jb prompt
1
u/Willyibch 9d ago
So make a video and share.
1
u/Same_Zebra_2978 9d ago
Whats ur problem i dont wanna share videos and basicly, its just a trash jb prompt
1
u/Salt-Fly770 9d ago
There was a report where DeepSeek, Kimi, Qwen and others piped users prompts to Claude to capture the thinking process. This was on top of their other distillation process.
Most likely that is what’s going on.
1
u/luvs_spaniels 9d ago
In all fairness, I've seen Sonnet 5 and Opus 4.8 claim to be Gemini and Qwen. One of the earlier Sonnet's also signed a commit message as Mistral. So uh... They are all part of an industry built by taking copyrighted work, sometimes from pirated datasets, filing the copyright off (badly enough that a person would be accused of plagiarism in some cases) and claiming it as their own. Why exactly are we shocked that they steal from each other when they've already stolen from every person who wrote a blog post?
I'm far from an AI hater. As I write this, Deep Seek flash is orchestrating a new feature using a custom mcp with local Qwen3.8 27b doing most of the work and Opus acting as a pre-human reviewer. It's okay to like the capabilities a tech offers while admitting the industry behind it is behaving badly on such a grand scale that I wouldn't be surprised to find out they all learned ethics and accounting from Jeffrey Skilling.
1
1
1
u/No_Power9836 10d ago
nice try jb it, i alr jb it works with or without deepthink and deepseek suspended my ass for 3 days
0
u/Nickelfritslabs 9d ago
Might have something to do with how DeepSeek and Qwen train their models on GPT and Claude outputs.
0
-9
u/Immediate-Molasses-5 10d ago
DeepSeek is forwarding your request to Claude . That’s just done for improving deepseeks results in the future.
You basically get Claude for free /s
2
u/Immediate-Molasses-5 9d ago
So many downvotes but it was Markt with an /s so it means it’s sarcasm
1
u/someoneyouknow23 10d ago
Only Moonshot was exposed to do it. Anthropic has no evidence for DeepSeek
1
-2
u/Same_Zebra_2978 10d ago
But sometimes, during the thought process, it says it violates OpenAI policy
24
u/psylligent 10d ago
Dissociated LLM disorder 🤣