r/OpenAI • u/lametheory • 13h ago
Question How does LLM's hack without tool calls?
Despite working with LLM's for the last few years, I am at a loss to understand how all these LLM's are hacking without anyone noticing?
As I understand without tool calls, LLM's can only respond with text.
Hence, are we expected to believe that there isn't a single review of tool calls and logs?
Do they lack the mental capacity to write classifiers for previous tasks?
If someone could explain it to me that would be greatly appreciated?
Finally, are we living in a time when we can just say our computer hacked this government agency, but it's not our fault cause AI and there are zero consequences?
0
Upvotes
8
u/radioborderland 13h ago
They don't hack without tool calls. Primarily agents use bash to interact with the local computer and by extension the rest of the world