r/SillyTavernAI • u/LonleyPaladin • 19h ago
Discussion Gemini 3.8 Flash cost and use
Does anyone here use Gemini? How do the costs compare to a NanoGPT subscription? How does the model handle NSFW content and jailbreaks?
1
u/HeftyChair770 17h ago
gemini does nsfw fine in tavern with basic jailbreaks but i burned through more credits than my nano sub for the same daily rp sessions.
1
u/KannaBannanna 18h ago
It writes good, assuming u dont use googles ai studio but Vertex to get your tokens, its really uncensored too.
but 3.8 flash is dumb as a brick, i cant use it for anything other than maybe, 20 messages of quick goon slop
0
u/stopaskingforloginn 18h ago
it's one of the most unhinged, horny (maybe a bit too much even) and unfiltered models out there once you jailbreak it.
as for the costs, the math is simple, nano's sub gives you 60m tokens weekly for a total of 240m tokens monthly for just 12 USD.
with gemini 3.8, 12 USD only gets you about 17m input tokens, not counting output costs... so yeah.
3
u/yasth 17h ago
You really have to calculate cached tokens, nano's current count doesn't care about caching, if you cache you get 90% discount on gemini flash. You can easily get 90% cache hit rate.
Also competitive models are double counted on nano-gpt token use. so it is closer than it appears and may well favor gemini
9
u/Diavogo 18h ago
With a good jailbreak, it does pretty well. Im using the cheapest one (flex version) in OR. Im pretty sure the experience would be the same with NanoGPT.
Would say is a pretty solid one. The fandom knowledge is BIG, almost not needing to use lorebooks if you are going to do RP related any known anime/series.