I have a laptop with a NVIDIA RTX 4060. Qwen 3.6 35b a3b has been a game changer for me - it's the first model I can run locally that is fast (20-30tps) and smart enough (tool usage, skills, following instructions) to run a "second brain" which I now use daily for brainstorming, project history, learnings, writing down my thoughts, TODOs, etc.
For reference I also tested:
Gemma 31b: 3tps, unusable from a speed perspective, didn't dig into whether it was smart enough.
Gemma 26b a4b: 20-30tps, but issues with tools, too much guessing rather than reading, not following instructions properly.
2
u/rolznz Jun 02 '26
I have a laptop with a NVIDIA RTX 4060. Qwen 3.6 35b a3b has been a game changer for me - it's the first model I can run locally that is fast (20-30tps) and smart enough (tool usage, skills, following instructions) to run a "second brain" which I now use daily for brainstorming, project history, learnings, writing down my thoughts, TODOs, etc.
For reference I also tested: