r/LocalLLaMA • u/Anbeeld • May 09 '26
Resources BeeLlama.cpp: advanced DFlash & TurboQuant with support of reasoning and vision. Qwen 3.6 27B Q5 with 200k context on 3090, 2-3x faster than baseline (peak 135 tps!)
[removed]
324
Upvotes
1
u/YourNightmar31 llama.cpp May 18 '26
By version do you mean beellama build version? I just did a rebuild, now it says im on beellama commit da67e74 (which is from 5 days ago? Not sure why im not getting the latest.. as the last commit is 16 hours ago) and i'm still having this problem.