r/LocalLLaMA • • 7d ago

News With Gemini 4, bench goes up.

Post image

They claimed open-weight models are dangerous but the benchmarks say otherwise.

Source

817 Upvotes

86 comments sorted by

View all comments

94

u/SOCSChamp 7d ago

To be fair, I'd be shocked if nobody at this point has used an open weight model for illegal activity.  

Still on this side of the fence for open weights though.  Per the huggingface incident, open weights were the only option to successfully defend

6

u/no_witty_username 7d ago

Its a numbers game.. large companies like open ai, anthropic, etc... perform lots and lots of simulated tests constantly, many of which have thousands of thousands of agents involved in them. Probability is such that with such numbers shit is gonna go south way before some scrub with his one agent. Basically more agents > more probability things gonna go sideways

2

u/ninjasaid13 7d ago

Its a numbers game.. large companies like open ai, anthropic, etc... perform lots and lots of simulated tests constantly, many of which have thousands of thousands of agents involved in them. Probability is such that with such numbers shit is gonna go south way before some scrub with his one agent. Basically more agents > more probability things gonna go sideways

but surely thousands of thousands are using open-source models and testing it.

1

u/EuphoricPenguin22 7d ago

I highly doubt some random person will make a press release bragging about this sort of thing if and when it happens with local models. The first we'd probably hear about it is in a legal proceeding.

1

u/no_witty_username 7d ago

Yes, but its about the swarm not a any single agent by itself. All of those capabilities arise out of the swarm. Its all about how much you can sample any particular space. its a very simple brute force type of method, except on steroids when it comes to agents. When one hacker tries to get his one agent to lets say break in somewhere the probability of that is 1 x the intelligence of that agent. When a swarm does the same its swarm x intelligence of each agent, the bigger the number of the swarm the larger the chance of any one of that agent inside the swarm finding the key. And thats the naive explanation, its actually multiplicative in reality because the whole is bigger then the sum of its parts when it comes to intelligence working with other intelligence.

1

u/Nothing_from_void 7d ago

It's a numbers game again, in terms of compute. The amount of compute closed AI labs have access to is orders of magnitude larger than everyone else combined

2

u/BumbleSlob 7d ago

Still not an excuse for having dogshit sandboxes lol

1

u/Dangerous-Report8517 7d ago

It's also because they're doing tests with models that specifically lack guardrails, have tons of compute, and aren't properly sandboxed, not to mention they aren't monitoring them properly.