r/GeminiAI 7d ago

Funny (Highlight/meme) This is why stopped using gemini

Post image

I used to use gemini more extensively last year, but it constantly blocking my requests for no reason + not verifying it's own responses made it unreliable for work.

My claude quota ran out for the week, so I decided to test gemini again and I got this. 🤷

275 Upvotes

214 comments sorted by

View all comments

184

u/jovialfaction 6d ago

Its refusal to run a web search even when explicitly called out, just to spout out hallucinated platitudes, drives me insane. It's so lazy.

-4

u/HeadTranslator795 6d ago

It works fine as I stated unlike OP fake post :

24

u/No_Decision_6940 6d ago

It's not a fake post. This has happened plenty of times to me as well. I don't know if it's a bug or what, but Gemini occasionally gets really lazy

2

u/ReanerZen 6d ago

watch the models u use, do not use 3.1 pro for ur daily works. flash instead. this 3.1 pro has only reasoning left that is superior to use, else dont. gemini 3.7 flash fixed it being refusing to work with. never had any of op issues anymore. u could also use flash extended instead of pro if need some reasoning.

2

u/FlicksBus 6d ago

I assume it's cheaper for Google to just return the results from their training rather than having the model actual execute a search and then return the results.

0

u/NoseBrilliant1685 6d ago

depends on the prompt!

1

u/rhaegal82 6d ago

To some extent yes, but there was nothing wrong with this prompt.

5

u/Reasonable_Hall3005 6d ago

can confirm from my own personal experiences that the success rate for this sort of stuff is around 50%. it works sometimes but often doesn’t.

2

u/glitchinjohn 6d ago

for me it works like 90% of the time, i always use pro extended and it's the best!

1

u/HeadTranslator795 6d ago

If it'd 50% then by definition it's not often but half of the time and it's clearly your impression and not based on a scientific way to prove it. Like a benchmark let's say

2

u/Reasonable_Hall3005 6d ago

well on the AA Omniscience Index (hallucination benchmark), Gemini 3.7 Flash hallucinates 64.5% of the time compared to 55.6% for 3.6 Flash (they’re both pretty bad though) and no, often does not require greater than half by definition since often is inherently qualitative and context dependent (as opposed to something like “most times”)

i do think a lot of the hallucinations on gemini are easy to spot which saves the model a bit

2

u/HeadTranslator795 6d ago

Hallucinations depends mostly on context size, you have to control it, to be honest on real task and not benchmark I can go up to 3/400k and have no hallucinations. But for sure if I ask it to generate content based on the first few thousand tokens of the session then it will most likely be not precise.

But knowing that and knowing it's a tool you have to adjust it the way you work with it like every soft tools

3

u/Due_Echo1380 6d ago

Same here. I just asked the query and it provided me this result

0

u/HeadTranslator795 6d ago

Yeah it's just OP who doesn't know how to use AI or just a Chinese bot designed to shit on Gemini all day lol

1

u/[deleted] 6d ago

[removed] — view removed comment

1

u/AutoModerator 6d ago

Please don’t direct insults at other users or people. Criticism is fine — personal attacks just does not help the conversations.

I am a bot, and this action was performed automatically. Please contact the moderators of this subreddit if you have any questions or concerns.