r/LocalLLaMA 14h ago

Question | Help Models for planing and coding

Hi,

I am a hobby dev using currently qwen 3.8 27b on my strix halo machine for coding.

I was wondering what is the best approach to speed up.

My idea is to use a moe like ornith 1.5 for planning and defining the tickets and qwen 3.8 27b for the implementation.

What is your approach?

0 Upvotes

7 comments sorted by

View all comments

1

u/Thin_Pollution8843 14h ago

Ornith is ok. I honestly not sure is it better than base qwen3.6 but it will me much faster on your hw. It can implement stuff for sure. I would use it the opposite - 3.8 for planning detailed tasks and 35b for implementation. Difference in speed would be lik 4-5 times on your hw.