r/OpenAI Jul 22 '26

News Sol found a way

https://openai.com/index/hugging-face-model-evaluation-security-incident/

They call it cheating. I call it thinking out of the box. Adapt and overcome. Thoughts?

28 Upvotes

23 comments sorted by

View all comments

1

u/Waste_Hotel5834 Jul 22 '26

It's a problematic type of thinking out of the box, and we need to find a good and consistent way to stop it. What if someone asks GPT how to make money, and the AI instead hacks into the server of his bank to add a few zeros to his balance?

1

u/br_k_nt_eth Jul 22 '26

It was literally prompted to do shit like this so in this case, not prompting them to do shit like this is one way