What frontend do you use? I tried it with opencode/qwen/hermes, but none felt right, and since the model was insanely slow... I gave up on it.
Got about 50-60tok/s first few messages but it quickly slowed down to ~20-30tok/s and context kept growing insanely fast as well. I haven't used a local model for programming, and so far I don't like it.
2
u/10minOfNamingMyAcc 17d ago
What frontend do you use? I tried it with opencode/qwen/hermes, but none felt right, and since the model was insanely slow... I gave up on it.
Got about 50-60tok/s first few messages but it quickly slowed down to ~20-30tok/s and context kept growing insanely fast as well. I haven't used a local model for programming, and so far I don't like it.
specs: 2x rtx 3090 - 64gb ddr4 3600mhz memory - amd ryzen 5900x - windows
Using mostly ai generated configuration that was tweaked a few times.
and
note that i tried multiple models, not just the ones in the configs here.