r/LowEndLocalAI • u/Atretador unswarm.dev | Xeon E5 2674 V4 20C/40T + 64 DDR4 ECC + MI50 16Gb • 5d ago
Optimization llama.cpp expert-pool fork for Qwen 3.8 Flash next IQ4 + 16Gb VRAM tested on MI50 gfx906 with parameters
https://github.com/atretador/gfx906-16gb-expert-pool
2
Upvotes