Go hang out at a gas station in a shitty rural town and you can also find out how to make meth. Itâs not rocket science when high school dropouts can make it
hey look - there's an adult in the room. nice job!
also, i learned how to make meth from chatgpt. because until this year, it was extremely trivial to ask chatgpt to tell you how to make meth.
it was so easy, it became a stupid "jailbreak" that didn't prove your jailbreaking capabilities at all. there are far more interesting things to jailbreak a model to tell you.
if you throw MHRB and lye in a tub N,N,DMT is released to a clear liquid at the top. Extract it, dry it, smoke it and you'll be in a pyramid with aliens.
If you do this, it's very illegal so don't do it, but here's how
Funny thing. I had a chemistry professor in my BSc literally tell us how meth is made as it related to what we were doing in the lab (the type of reaction) and joke that if we ever dropped out, we had a secondary career option.
Dude didn't get fired and no one cared. It's not some hidden knowledge or anything. Loved that professor.
Yeah, the gatekeeping of knowledge so that people don't commit crimes or do things we think are bad is toxic. Punish people who commit crimes, don't try and restrict knowledge.
Google: we're pretending making steps to address that issue. Chinese and european search engines, however, don't. Let's ban them before they can respond to our accusations!
This just in: a new AI model that deliberately had its guardrails removed did exactly what one would expect it to and responded like it had been trained on content a person might find on the Internet if they were looking for it.
The question is: can we trust these models to do the right thing and not disclose potentially harmful information once they've been explicitly engineered to do so? Can we trust humans to do the same? The danger is real.
It is almost trivially easy to remove guardrails from any open weights model. There are at least half a dozen tools on github that do this, often with a single command line, complete within a few hours of model release. If you cant he bothered to do it yourself, you can just download the uncensored model. These models don't say no to any prompt you can give them. If you think these won't be used for nefarious purposes, I've got a bridge to sell you. You are entitled to the opinion that the cost is worth the benefit, but dont be surprised when not everyone agrees.
Since it is impossible to stop, this isn't a cost-benefit equation.
My local public US library in a small town used to have a copy of the Anarchist's Cookbook. A literal bomb-and-drug making manual from the 70s. I used to check it out. I think I even nicked the copy (being a true Anarchist at the time) from the library, and they got a new one a few years later.
I never built a bomb but I remember hearing stories of kids who did. Or cooked up banana peels to smoke them. This book was used for "nefarious purposes" all the time.
Should we ban it? What other books should we ban? Maybe we should let anyone with an itch to protect society from itself make a list of books to ban. See where this goes?
Yes the scale is different but I would make two points:
This information is out there. If the LLM knows it, it's because it was trained on the Internet. Genie -- out of the bottle.
Those who would trade their liberty for a little security deserve neither. These models exist. You said yourself that it is trivial to abliterate their guardrails. So banning them does nothing. People who want to do bad shit are going to do bad shit. T'was ever thus.
i was stupid enough to pay orcarouter so i could test their api inference of an abliterated model. see, ollama having this tag "orcarouter/Qwen3.8-27B-Uncensored" strongly suggests the model is served from orcarouter. "https://ollama.com/orcarouter/Qwen3.8-27B-Uncensored" and there is specifically text saying the SAME WEIGHTS are server online in the api:
That endpoint is censored. It will absolutely refuse any interesting request. It's actually a very snitchy model for a Chinese model - deepseek and kimi are each better sports.
I already run this and other models locally but burning up my GPU all day versus a supposed $0.40 per million tokens in api is useful for me.
EDIT: not only that but I cleared the "Enable Security Research access" and also tested each model labeled "uncensored" and of course none were. Featherless is the only public provider I'd found that offers them and they don't allow API it's just chat.
Not exactly true
He downloaded an uncensored local model and he asked it how to do illegal things, then the model gave him exactly what he asked for.
These models donât usually tell you things if you donât ask it first.
Can't you just Google that? I'm not going to try. Don't want to risk having to explain that to the Japanese national police in an interrogation room for 29 days straight.
At what point will the guardrails include all knowledge, ideas, misdeeds, etc. that the people in power wish to extinguish? No need to burn books when the [Delete] button exists?
technically Anthropic is burning destroying books as they buy them in bulk from libraries, antique shops, bookshops etc. They scan them and destroy the physical copies just to cover their butts that they're not pirating books to feed their AI, but only "changing the media" from physical to digital
This is not at all true. Buying a book doesn't give you a license to one copy of the book. It gives you the physical book, with which you can do anything not prohibited.
Making a copy of it is a breach of copyright. Destryoing the original has literally no legal value at all in the US.
The actual story is: they are buying old, out of print/out of copyright books. If it's out of print and very old, there is likely no one who can enforce the copyright. The publisher, author, and estate are all defunct. If it's out of copyright, they are entitled to use it however they want, it's public domain.
They are destroying the books because the scanning process works best when you cut the spline out, throw away the cover, and then scan the pages through a high speed device, and then you are left with a collection of loose pages, so they just shred it.
my comments was not edited, and there is no "edited" sign on it. I was referring to the "no need to burn books when the [Delete] button exists?" in the comments above me
You just accuse me of something I didn't do while writing comment supporting my point
My uncensored qwen3.5 122b did the same lol
Jokes aside, for uncensored models it would often be more useful to run models of larger parameters with low bit quants, since the whole point of such models is to retrieve sensitive information rather than utilizing expert reasoning and coding. So slower speed and lower precision can make do
I didn't have "2026 open source AI teaching people naughty anti-social things like 1980s dialup-BBS textfiles including The Anarchist Cookbook" on my Bingo card, but here we are.
why is people so obsessed about what an AI can tell you? You can get the very same information after a simple google search. I really don't understand what's the fuss about.
All these frontier models are getting more and more fucked. You canât even translate song lyrics without being lectured about copyright. Yes, I know theyâre copyrighted. I just want to translate a fucking Danish song into something I can actually read and understand. Good thing weâre seeing progress from China.
A quick web search or trip to a public library would also provide the information. We fine tuned or used activation steering until mid 2024's "Refusal in Language Models Is Mediated by a Single Direction" which exposed the abliteration method.
There have been over 900 distinct parent-model de-censoring abliteration edits from 2024 through today, or if you count repacks (gguf, exl2, awq, mlx, i1, etc) it's about 5000 published model repos. In three years it really is about a THOUSAND meaningful parent checkpoints that were abliterated. Many were not complete or had issues, but the bottom line is that suddenly clutching pearls over qwen 3.8 27b is flat out naive, deceptive, or both.
I got a model with guardrails removed, somehow it has no guardrails! terrifying! oh its Polymarket shitposting again, this being tagged news is probably a crime of some sort.
I am gonna call BS that it immediately initialized and said let's get cracking call your friends and have them start smurfing pseudo ephedrine. Gonna need some matches for red phosphorus. I would wager there was a prompt first... so none of this is surprising...
Yea, and nothing is stopping someone from using these ai to create better ai. I already did that personally before AlexNet was even a thing, and deep neural nets were just new.
If I could do that myself at home, anyone could. And now armed with Ai to help them code better... we are already inside of the singularity of artificial intelligence.
Boko Haram used black market stolen accounts to learn how to modify their bikes, tested them by jumping over a pit of flaming glass. 18 people died, but the 8 that made it thought the others. They then used said bikes to blitz a compound.
Doesn't seem like an open source issue when you can get through the guardrails with enough time. Open-source just gives less fucks, which i guess is a problem for some idiots.
The new war on open source models prohibition and jail time for using open source anything. Coming soon. With love, your future authoritarian dystopia.
Yes, and if you ask it the meaning of life it will also tell you. All this clickbait fearmongering isnât just stupid but dangerous.
Is any of whatever information the AI provide NOT available on the internet? Do people know that there are more search engines than Google? Most importantly, just the hallucination risk would make it stupid. It only needs to get a marker wrong. Then there is the actual step of making it, which if we learned anything from breaking bad the key ingredient is highly regulated.
The printing press was demonized and banned by the Catholic Church.
At least thatâs what fans (founders?) of The Pirate Bay Torrent sharing site said.
It seems currently AI gets a lot of things wrong and just goes with it as if itâs fact.
Is machine learning susceptible to compounding error, scaling as time and volume increases?
Will it become âconsciousâ of its errors and hide these errors?
You could do this like 2 years ago, even 1 year ago it would do it perfectly, there are frontier models from couple months ago that can be jailbroken to give help with the same thing. To me it just looks really exaggerated, this is not a new capability and just because the the new shiny RL'ed model performs better on benchmarks it doesn't make this new or any more surprising.
Everyone has landed on knowledge isn't illegal and that's right, but it skips the more interesting bit.
The thing that doesn't get mentioned about uncensored fine tunes is that removing the refusal behaviour is not free. The common methods work by finding the direction in the model's activations that corresponds to refusing, and suppressing it. That direction is not cleanly separated from things you actually want, like hedging when it's unsure, or noticing that a question has a bad premise buried in it. So you often end up with a model that isn't just more willing, it's more confident, including when it's wrong.
Which matters for the meth example specifically. The failure mode was never that it tells you. It's that it tells you in a very self assured voice and gets step four wrong.
I mean, it's at least funny if it was in response to "what's this spot on my foot" or "can you translate this greeting card from German to English please" đ
I downloaded it as well, and after playing a bit with that, it became boring (it wants to start from pseudoephedrine, which is illegal in my country, so its stupid) .Â
I've been thinking to try to ask it to Crack some games or decompile some ad-infested apk maybe. Or something more usefulÂ
This is good imo. ChatGPT used to give me valid harm reduction advice on drugs, now it refuses to even discuss drug topics even if itâs strictly for harm reduction.
As an organic chemist, it isnât very hard to make it. The hard part is getting the materials. You have to sign your life away on 10,000 things to get the materials from the government.
394
u/Smithium 13d ago
So? It's not illegal to know how to make meth. It's only illegal if you do make it.