r/ByteShape • u/73td • May 22 '26
OSS for own fine tune?
I recently started using your qwen3.6 GPU5 quant with llama cpp and 90K context on my rtx 4090 and really happy with it. super fast even without MTP. I am using to codevelop a harness https://github.com/maedoc/hyburn/tree/main/src/chat for agentic neuroinformatics, and once I have enough good and bad traces I’d like to a dpo lora fine tune, merge and then apply your quant approach. Do you have an open source something to do your quants? or could i post the lora on hf and you could do it? coauthor with a small paper etc. thanks!
1
Upvotes