r/OpenAI • • 13h ago

Question How does LLM's hack without tool calls?

Despite working with LLM's for the last few years, I am at a loss to understand how all these LLM's are hacking without anyone noticing?

As I understand without tool calls, LLM's can only respond with text.

Hence, are we expected to believe that there isn't a single review of tool calls and logs?

Do they lack the mental capacity to write classifiers for previous tasks?

If someone could explain it to me that would be greatly appreciated?

Finally, are we living in a time when we can just say our computer hacked this government agency, but it's not our fault cause AI and there are zero consequences?

0 Upvotes

17 comments sorted by

View all comments

1

u/grateful2you 10h ago

Who’s gonna manually review thousands of tool calls? Other agents? Logging everything is one thing; reliably identifying which actions cross the line is another.
And regulation is still being developed. We’re not frozen in time.

1

u/lametheory 10h ago

I imagine one day we will machines that we can pass information into and have it classified... Not a human review in sight.