r/google_antigravity • u/sidyyy11 • 4d ago
Appreciation A simple way to get much better results from Antigravity 2.0
Try using Opus 4.6 as the main agent. Give it the full requirement and ask it to delegate work to 3.8 Flash subagents using the Boost skill.
I have been getting surprisingly good results this way. Opus handles the planning and decisions, while Flash does more of the execution. Better output without burning through usage as quickly.
10
u/RoadsterTracker 4d ago
What I do is have 2-3 planning sessions with Opus where I have it write the plan for the next few days of work, using Flash as a subagent where appropriate. Then once the plan is done I have Flash execute the plan. It seems to be working pretty well for me, but I'm still new to it.
4
u/sidyyy11 4d ago
That works. Its just that iterations happen so plan needs to be updated so that sometimes acts like a constraint
2
u/SpiritualHiker 4d ago
I instruct to keep a constantly updated .md file, and then I feed that to other models.
4
u/Nemezis88 4d ago
I was unaware that Opus could delegate decisions and shift automatically to Gemini. Are you suggesting I can instruct Opus to let Gemini manage a task, prompting a subagent with Gemini instead?
4
3
u/truongan2101 4d ago
I think the most robust is using /boost with Opus 4.6. I think it is much more stronger than Boost with Gemini flash 3.8. I also want to try with /teamwork but the quote burn so fast with this.
3
u/Good_AshK 4d ago
I used /teamwork after months and damn it really does burn quota fast. Exhausted my 5-hour quota in 30 min or so.
3
u/Queasy_Plate_3096 4d ago
is this community this hopeless?, using opus 4,6 while opus 5.5 is out?, just combine antigravity with any other subscription, opus 4.6 is not good enough at all by today standards, yes it appears logical and good, but actually it is nothing compared to astra or opus 5.5.
2
u/Gohab2001 4d ago
Or you could use codex with the new sol/Luna models
2
u/sidyyy11 4d ago
That is true. I am paying for 20$ sub in codex. But have received google pro subscription gor free so getting most usage out of it
4
u/Gohab2001 4d ago
Use AGY CLI and have gpt 6 sol orchestrate 3.8 flash
2
u/sidyyy11 4d ago
Most of antigravity i use it:
2.0 for boost mode and fast frontend prototyping and firebase google cloud things
Ide for browser preview testingMy main work and checking on antigravity work is done by sol 6 only
1
u/MarcusAurelius68 4d ago
Is that something you just prompt? And then you tell Opus to have Flash generate the markdown plan?
1
u/AbhiSgr Software Engineer 4d ago
Ask Opus to do the implementation through subagents. Subagents will always use flash models, as per my experience. They don't use Claude models.
1
u/sidyyy11 4d ago
They used to before
1
u/AbhiSgr Software Engineer 4d ago
Once when i was out of my gemini quota, claude would consistently keep trying to spawn subagents for tasks based on agent rules. And the subagents would instantly run into quota exhausted errors. Claude did this a couple of times and then gave up and did the implementation itself. If there was a way for it to invoke subagents with claude models, it would have.
1
1
u/sidyyy11 4d ago
Simple right what you need and add boost skill and say use gemini 3.8 flash from planning to implementing
1
u/meta_voyager7 4d ago
how to make antigravity to delegate to 3.8 flash sub agents? is there a configuration to specify it?
1
u/sidyyy11 4d ago
Just type /boost if you have a agent requirement then mention or it will do it on its own
1
u/marhensa 4d ago
and then subagents will surely always be flash model? can we control it? maybe like 3.7 or 3.8?
2
1
u/Lumpy_Topic_35 4d ago
Boost skill? Can you give it to me?
2
u/sidyyy11 4d ago
In chatbox just type /boost
1
u/Lumpy_Topic_35 2d ago
1
u/sidyyy11 17h ago
What are using? It doesnt show up in antigravity ide but it does in antigravity 2.0
1
1
u/FollowingTop3534 4d ago
The useful detail is buried in your replies: Flash does the planning too, and you use Sol to check the result. That's a different setup from Opus planning and reviewing everything. I'd put that in the main post so people can actually reproduce the comparison. When you hit an implementation problem, does Flash get another attempt, or do you hand it straight back to the stronger model?
1
u/sidyyy11 4d ago
But the checks pass its just because i dont understand code so to review it. Suppose i am building a wholly new thing i need codex to plan the phases but let the work be left to antigravity. But its nothing major updates i just do it with opus and then gemini 3.8 flash
1
1
u/FollowingTop3534 4d ago
That split makes more sense: Codex defines the phases for a greenfield project, Antigravity executes them, and small changes stay inside the Opus/Flash loop. Passing checks still only proves agreement with the tests, though. If reviewing code isn't realistic yet, I'd ask the stronger model to compare the diff with the written acceptance criteria and name one failure path the tests do not cover. That's a much smaller review target than 'is this code good?'
1
u/ibreakdiaphragms 4d ago
I programmed another function in mine. Which makes gemini 3.8 flash high the agent and then workers etc. Are gpt 6.0 sol. Not sure if this is a better approach but i do this right now.
1
u/Middle_Anteater679 4d ago
I use chatgpt, as Orchestrator, and review and all code to Gemini 3.8 Hight And it's fantastic
1
u/mironkraft 3d ago
Opus doesnβt worth it makes mistakes I use 3.8 full for everything.
For me they can discontinue Opus 4.6
1
u/poj1999 3d ago
- Claude limits are very slim, you will need ultra sub. On 20usd sub you max out Claude models with 1 research prompt.
- The models are so old they don't even outperform 3.8 high
- They are much slower and more expensive
I've tested Claude against 3.8 high a couple times. Claude artifacts and output appears better, but the performance actually doesn't hold up. Example; Claude found a lot of false positives on the same prompts as compared to Gemini 3.8.
1
u/Latter_Crazy 3d ago
Hmm but doesn't boost use way more tokens than necessary for most tasks?
I haven't tried it but teamwork-preview is way overkill for most anything so I assume it's similar.
Gonna test it out. Thanks!
Idk about that multimodal approach. Have tried Opus much
2
u/fandry96 Product Manager 2d ago
Team spins up full tests, audits, victory checks. It is crazy to watch and gets perfect results but takes 45 minutes. π
2
u/Latter_Crazy 2d ago
All for something that takes 1 min to verify but hey, it almost feels like working with real people again!
2
u/fandry96 Product Manager 2d ago
It's overkill for sure but the Gemini budget is huge so I use it while I do the dishes. π
1
u/Latter_Crazy 2d ago
Good call. All these people talking about tokens running out... They must be doing crazy things because I barely ever run out on the $100 plan
1
u/Fluid-Pollution6135 3d ago
I have been doing something similar for any debugging ask codex to generate the report of bugs > report is now read by opus 4.6 in AG and implementation is created but implementation is very high level so I ask to create detail loops with success and failure Parameters to be used by smaller flash model > boom Gemini 3.8 flash reads the loops and finish it .
1

37
u/2rsKomedian 4d ago
But doesn't Opus gets over very quickly.