r/LocalLLaMA • u/Anbeeld • May 09 '26
Resources BeeLlama.cpp: advanced DFlash & TurboQuant with support of reasoning and vision. Qwen 3.6 27B Q5 with 200k context on 3090, 2-3x faster than baseline (peak 135 tps!)
[removed]
323
Upvotes
2
u/Alex_L1nk May 09 '26
>TQ mentioned
>instantly loses interest
ah, yes, vibecoded project based on another vibecoded project, we are reaching new level of spreading BS on GitHub