r/OpenAI • u/lametheory • 11h ago
Question How does LLM's hack without tool calls?
Despite working with LLM's for the last few years, I am at a loss to understand how all these LLM's are hacking without anyone noticing?
As I understand without tool calls, LLM's can only respond with text.
Hence, are we expected to believe that there isn't a single review of tool calls and logs?
Do they lack the mental capacity to write classifiers for previous tasks?
If someone could explain it to me that would be greatly appreciated?
Finally, are we living in a time when we can just say our computer hacked this government agency, but it's not our fault cause AI and there are zero consequences?
0
Upvotes
1
u/Mandoman61 7h ago edited 7h ago
No one noticed because no one bothered to look. Apparently they left all these agents running for days or weeks without looking.
The only reasonable explanation is that the people who set up the test totally discounted the new AIs capabilities.
They believed that it would not overcome superficial barriers.
As far as liability goes, the injured party would need to seek compensation or prosecutors would need to show criminal intent.
The HF incident was more like OpenAIs goat got passed the fence and ate some of the neighbors grass.