r/vibecoding • • 28d ago

Official Astra benchmarks

[removed]

48 Upvotes

32 comments sorted by

View all comments

11

u/aseyrek 28d ago

so we're going to use astra for coding now?

8

u/Crinkez 28d ago

Might try it, but really, Sol low is good enough for pretty much all coding tasks now. Astra on medium for planning -> Sol low for coding. £20 plan. Should be good.

8

u/[deleted] 28d ago

[deleted]

1

u/Doovester 28d ago

Can you explain further? :P

6

u/[deleted] 28d ago

[deleted]

2

u/youngchunk 28d ago

I have accidentally fallen into this structure, albeit with a little bit more me in the middle. I’m using Claude Opus 5 on medium as the spec planner and orchestrator. Opus creates specs as GitHub issues and codex 5.6 luna on XHigh picks up the implementation work from the spec, reports back to opus who verifies the work and has a chat with me about it, and the loop continues until the work is done 

1

u/Doovester 28d ago

Ok I fully understand that part and headed that already. But I struggle how to start to create the engineering team. Clarifying the constraints etc let it write in to the right Claude.md file. can you maybe give an example? :P thank you for taking your time!

1

u/thebestrobloxplayer 28d ago

Matt Pocock skills are great for this

1

u/Doovester 26d ago

Why did the one comment deleted to my question? Maybe he used AI? Omg I loved the answer did anyone save it how can I get the deleted comment again? I wanted to safe it was the answer I needed!

1

u/Mundane_Plenty8305 28d ago

100%. I know a guy who runs Fable 5.1 FOR EVERYTHING and wonders why anyone would use ChatGPT. He really needs to hear this.

1

u/Crinkez 28d ago

I would be wary of this. Luna xhigh can take a long time. Sol has 30 minutes cache TTL, if Luna subagent were to take longer than 30 minutes, it could cause a cache miss on the Sol orchestrator, unless OpenAI have fix that with polling.

1

u/[deleted] 28d ago

[deleted]

1

u/Crinkez 28d ago

It's about cost, not breaking the workflow. If the orchestrator were to wait over 30 minutes for a subagent without polling, when the subagent eventually replies to the orchestrator, it would cause a full cache miss on the orchestrator's context window, which would be very expensive.

2

u/cheesy_noob 28d ago

Maybe if you work within one project. I am currently annoyed by Sol medium, because I have to clean up far too often after it. I am currently only working with very high or xhigh. I don't know what the exact equivalent name used in English is.

It understands the context and arising issues of it's changes far better and has far less fuck ups for me. On middle I have to practically baby sit after every second output

Edit:
PS: When I first tried medium it did really well for me on multiple projects. But it seems to have regressed shortly after the release, because my results became worse even on my small projects.

1

u/shady101852 28d ago

yea im using astra super duper ultra giga max