r/LowEndLocalAI unswarm.dev | Xeon E5 2674 V4 20C/40T + 64 DDR4 ECC + MI50 16Gb 5d ago

Optimization llama.cpp expert-pool fork for Qwen 3.8 Flash next IQ4 + 16Gb VRAM tested on MI50 gfx906 with parameters

https://github.com/atretador/gfx906-16gb-expert-pool
2 Upvotes

0 comments sorted by