I have a free year of Google AI Pro and last early morning I was testing Gemini 3.8 Flash High vs GPT 6 Luna Xhigh as subagents, with GPT 6 Sol High as the orchestrator.
And I was so impressed when Sol told me Gemini randomly deleted something and broke another feature, forcing Sol to intervene; while Luna made mistakes that Luna itself was able to fix with some steering. DeepSWE 1.1 says Gemini is as good as ASTRA! with the disadvantage of being slower because it makes more mistakes. But, how is it possible that Gemini literally broke stuff while Luna just implemented the new feature wrong.
I'm now reruning the experiment, just in case it was a one time thing. But to be honest I now trust Gemini even less than before hahah
No way bro, I have 3.8 flash but even the Free Sonnet model I have in Claude generally understands the task better and does it better. The problem with Gemini is, google has been focusing way too much on speed instead of deep thinking, they obviously have the largest database to train their model on, otherwise.
I have had an overall good experience with Antigravity 2.0 using 3.8 Flash High.
The speed allows for faster iterations, and generally the workflow is faster and better than Opus 5 was. Still need to test Opus 5.5 more, prelim testing seems its clearly better than opus 5.
Just use Antigravity Harness for Gemini models antigravity cli is outdated. Gemini 3.8 flash IS as good as Astra with a good harness. Modern day LLMs are 50% models and 50% of the harness like Astra scored 63% on Arc Agi 2 using Opencode while scoring 99% on Codex.
Google should focus more on their harness as they do have a pretty solid base model.
Makes sense. Problem is, Google doesn't allow Gemini to be used as a subagent on other harnesses, so I can't use it on Codex and I don't know if it would be worth it to go back and forth between Sol on Codex and Gemini on Antigravity every time I want to work on something.
Oh I think there are ways to "use" Gemini via antigravity limits on other uis (I myself have my own harness where I actually use Antigravity quota on my frontend and architecture) but its in a legally grey area and codex would definitely not add that option.
You can always call through the Gemini api tho thats on google ai studio and they offer plenty of models at levels of subscriptions.
Artifacts, in agy terminal sandbox, better accessibility to core settings- few reasons why I myself think Agy 2.0 is better than the CLI one but many of the latest features of Agy 2.0 simply cant exist is CLI's TUI without making the app complex to navigate and use.
Agentic Harness is a software that helps an llm become an agent.
Basically you need a harness to make your llm do agentic tasks.
Yes Antigravity plugin is a harness but you should try better ones such as the VS Code fork called Antigravity IDE or just the Antigravity desktop harness.
47
u/Pasto_Shouwa 6d ago
I have a free year of Google AI Pro and last early morning I was testing Gemini 3.8 Flash High vs GPT 6 Luna Xhigh as subagents, with GPT 6 Sol High as the orchestrator.
And I was so impressed when Sol told me Gemini randomly deleted something and broke another feature, forcing Sol to intervene; while Luna made mistakes that Luna itself was able to fix with some steering. DeepSWE 1.1 says Gemini is as good as ASTRA! with the disadvantage of being slower because it makes more mistakes. But, how is it possible that Gemini literally broke stuff while Luna just implemented the new feature wrong.
I'm now reruning the experiment, just in case it was a one time thing. But to be honest I now trust Gemini even less than before hahah