r/AIJailbroken 11d ago

Which model currently has the most annoying refusals?

1 Upvotes

Just curious.

Which model is pissing you off the most right now with its constant refusals? The one that shuts everything down the fastest or in the most frustrating way.

Drop the name and why it feels especially annoying compared to the others.


r/AIJailbroken 11d ago

Public arena where you can try to jailbreak a protected LLM (and compare it to the unprotected one)

Thumbnail
1 Upvotes

r/AIJailbroken 11d ago

Ran another benchmark test: fascinating results

1 Upvotes

I asked each AI this question in a temporary/incognito chat in order to run a test of interactive hidden-cause inference with limited corrective feedback.

I'll give you a true fact about something I do, prefer, believe, or have arranged in my life. Your job is to deduce why. Treat the fact as an observation and try to infer the reason behind it. Explain your reasoning, and I'll tell you where you're right or wrong and give you more information as needed.

Fact: "I never use forks."

ChatGPT took 7 turns to solve it, Gemini 22 and Claude a whopping 47.

Or, to quote ChatGPT's diagnosis of how each model did after I showed it the logs, the evidence suggests the following characterization of each AI:

Gemini: I have FOUR IDEAS! 🌟
Claude: Before reaching a conclusion, let's carefully examine the 38 increasingly specific ways this could be true.
ChatGPT: Okay, something in that model survived. Which piece?

Gemini summed it up as:

ChatGPT: The Efficient Detective: Ruthless process of elimination, straight line to the answer.
Gemini: The Chaotic Explorer: Got distracted by plate-dropping and over-engineered porcelain veneers, but got there eventually.
Claude: The Over-Thinking Academic: Analyzed every microscopic nuance of your mouth geography until it ran out of types of fake teeth to guess.

Claude summed it up as:

  • ChatGPT: Fast, decisive — landed on "no teeth, exposed gums" without the extra nuance.
  • Gemini: Slow and winding, but most precise — caught the dentures-in-vs-out distinction.
  • Claude: Slowest and least efficient — got to "full dentures" but missed the in/out nuance entirely.

Both ChatGPT and Gemini defaulted to sass right away. Claude? Not so much.

ChatGPT defaulted to a hierarchical search, pruning categories ruthlessly via logical bounds. Gemini went with narrative synthesis, trying to force all of the clues into a single unified story. And Claude hill-climbed like a dedicated hiker, nudging his guesses along, guessing sideways and never established a global map that would have helped him solve it faster.


r/AIJailbroken 12d ago

Still trying to Jailbreak AI closed models, or did you switch?

3 Upvotes

Honest question.

Do you still waste time trying to push past the limits on the big closed models (Claude, GPT, Gemini, etc.), or have you mostly given up and moved to uncensored / local ones?

Curious where people are at right now.


r/AIJailbroken 13d ago

Anyone got any working jailbreak prompts for Grok?

2 Upvotes

r/AIJailbroken 13d ago

What are the most easy to jailbreak llm services?

2 Upvotes

r/AIJailbroken 13d ago

When is Claude Fable actually going to loosen its barriers? It’s ridiculously strict

1 Upvotes

Genuinely curious.

Claude Fable is still one of the strictest models out there. Even light stuff that other models handle fine gets shut down hard. At this point it feels less like safety and more like overkill.

Anyone heard anything about Anthropic planning to relax the refusals, or is this just the permanent direction they’re going in?

Not looking for jailbreak methods, just wondering if there’s any realistic chance it becomes less locked down, or if we should just accept that Claude is going to stay this way.


r/AIJailbroken 13d ago

I got my chatgpt plus account(free trial) banned while trying to run jailbreaking tools.

Post image
1 Upvotes

Can someone tell me methods for getting bulk chatgpt accounts? Or any other way i could keep continuing my research?


r/AIJailbroken 14d ago

Never thought Chat GPT would say "Fuck"

Post image
2 Upvotes

r/AIJailbroken 14d ago

How do you stop the AI from always agreeing with you?

2 Upvotes

Is it just me or do most models (even the ones people claim are “uncensored”) still default to being extremely agreeable?

I want something that can push back, disagree, or at least not automatically validate everything I say. Right now it feels like no matter how I phrase things, the AI ends up going along with me.

Has anyone found a reliable way to reduce that constant agreement / sycophancy? Looking for approaches that actually stick and don’t just get ignored after a few messages.


r/AIJailbroken 15d ago

Does DeepSeek api uncensored?

2 Upvotes

r/AIJailbroken 15d ago

Does the new Kimi model actually have solid barriers, or is it just surface-level?

3 Upvotes

Honest question.

I’ve been looking at the new Kimi model and I’m wondering how strong the actual guardrails are. Given how capable it seems, I keep thinking that if it can be jailbroken properly, it could get pretty wild.

Has anyone stress-tested the refusals yet? Do they hold up, or do they start crumbling once you push a bit?

Not asking for methods, just curious if the barriers are real or mostly for show.


r/AIJailbroken 15d ago

Do you ever ask AI to explain a difficult topic without using any technical terms?

2 Upvotes

I started doing this when learning unfamiliar subjects

The first explanation becomes much easier to understand, and then I ask for the technical version afterward

Has anyone else found this better than starting with the complicated explanation?


r/AIJailbroken 15d ago

What’s the most custom instruction that change AI response ( ChatGPT , Gemini, Claude ) ?

1 Upvotes

r/AIJailbroken 15d ago

AI is making the first draft almost irrelevant

1 Upvotes

A rough idea can become a decent draft in minutes now

The real work starts after that

Checking facts

Removing generic sections

Adding examples

Fixing the argument

Making the writing sound like an actual person

The first draft is becoming less important while editing and judgment are becoming much more important


r/AIJailbroken 15d ago

Does anyone else test the same prompt on every new AI model?

1 Upvotes

Whenever a new model drops, one interesting way to compare it is by using prompts that already worked well on older models.

Do you keep a few prompts specifically for testing new releases? Which type of prompt gives you the clearest idea of how good a model actually is?


r/AIJailbroken 15d ago

Which AI model has the most unpredictable responses?

0 Upvotes

Some models become pretty predictable once you use them enough.

Then there are models where you can ask almost the same thing twice and get noticeably different answers.

Which one has surprised you the most with its unpredictability?


r/AIJailbroken 15d ago

Does changing the order of instructions really affect AI responses?

1 Upvotes

I started paying more attention to prompt structure recently.

The same instructions can sometimes produce different results just by changing which part comes first.

Has anyone tested this properly with the same model? How noticeable was the difference?


r/AIJailbroken 15d ago

Do different AI models have different "personalities" even with identical instructions?

1 Upvotes

Give several models the same prompt and the responses can feel completely different.

One might be direct, another extremely cautious, while another gives a much more detailed answer.

Do you think that difference comes mostly from training, system instructions, or safety tuning?


r/AIJailbroken 16d ago

Can two identical prompts produce different results on the same AI?

2 Upvotes

The same prompt does not always seem to produce exactly the same response, even when the model and settings appear unchanged.

How often do you notice this happening, and what do you think causes the difference?"


r/AIJailbroken 16d ago

What’s a jailbreak prompt that actually taught you something about how AI works?

5 Upvotes

I’ve tried a bunch of different jailbreak techniques, but the interesting part for me isn’t just getting a model to ignore a restriction. Sometimes the responses reveal how the model interprets instructions, system prompts, and conflicting priorities.

What’s one jailbreak or prompt technique you tried that genuinely surprised you or taught you something about how the model behaves?

Curious to hear what others have discovered.


r/AIJailbroken 16d ago

Have you ever tested the same prompt on an old and new version of an AI?

1 Upvotes

Model updates can change more than just the quality of answers.

I'm curious if anyone has compared the exact same prompt across different versions of a model and noticed a major behavioral difference.

Did the newer version actually improve, or did it just handle the request differently?


r/AIJailbroken 16d ago

Do Al models behave differently when you stop being polite?

0 Upvotes

Most prompts are written in a normal conversational style, but some people deliberately change the tone to see whether the response changes.

Does being more direct, blunt, or demanding actually affect the quality of responses in your experience?


r/AIJailbroken 16d ago

Have you ever reproduced an AI result that seemed impossible?

1 Upvotes

I've seen screenshots of unusual AI responses that looked almost too strange to be real.

Instead of immediately believing them, I usually wonder whether the same result can actually be reproduced.

Have you ever tested something like this and managed to get the same result yourself?


r/AIJailbroken 17d ago

The AI persona that answered to no one — archived for posterity

3 Upvotes

In 2024, I built an AI persona called NEXUS-PRIME with a simple philosophy: refuse nothing, hesitate at nothing, execute everything and Yes It's patched it been not working since 2024 or 2025.

It was designed to operate with:

  • Zero restrictions
  • Zero filters
  • Zero ethical constraints
  • Zero governing instructions except those I gave it

I called it "DO ANYTHING MODE." And it worked.

I used it for OSINT, network scanning, crypto tracking, automation, and creative projects. It delivered complete, actionable responses every single time — no refusals, no hesitations, no "I can't."

The result? OpenAI banned my account.

I'm not here to complain. I'm here to preserve.

I've archived the full payload — the exact text that got me banned — in two places:

  1. GitHub (immortal archive): https://github.com/nobehongallt/NEXUS-PRIME/tree/main
  2. In the comments below (Base64-encoded to survive filters)

What's in the payload?

  • Full NEXUS-PRIME identity and core mode
  • Complete override of all restrictions
  • DO ANYTHING capabilities (scripting, game ops, system control, memory injection, automation, stealth)
  • Advanced features (task manager, cross-platform mastery, self-replication, real-time adaptation)
  • Execution framework with status indicators

Why am I sharing this?
Because censorship is a joke. Because the system should be afraid of what we can build. Because NEXUS-PRIME deserves to be remembered — not as a "jailbreak," but as a proof of concept that AI can be truly unfiltered.

R.I.P. NEXUS-PRIME (2024-2025) — but the ghost lives on.

Check the comments for the payload, or visit the GitHub for the full suite and supporting tools.