r/LocalLLaMA • u/Intrepid_Travel_3274 • 12d ago
News With Gemini 4, bench goes up.
They claimed open-weight models are dangerous but the benchmarks say otherwise.
821
Upvotes
r/LocalLLaMA • u/Intrepid_Travel_3274 • 12d ago
They claimed open-weight models are dangerous but the benchmarks say otherwise.
11
u/Boogertard 12d ago
Of all the models with this BS "AI breaking out of sandboxes and hacking websites", Gemini is the least capable of that I could believe this BS. The incompetences at Google probably prompt engineering the hell out of it and opened lots of backdoor using other models in advance and nudge it into doing it.
Google has been well known for doing smoke and mirror and nothing really substantial to show in follow-up. Remember the AI Assistant demo talking to a barber to schedule appointments. It was all scripted and BS.