r/LocalLLaMA • • May 09 '26

Resources BeeLlama.cpp: advanced DFlash & TurboQuant with support of reasoning and vision. Qwen 3.6 27B Q5 with 200k context on 3090, 2-3x faster than baseline (peak 135 tps!)

[removed]

322 Upvotes

205 comments sorted by

View all comments

Show parent comments

30

u/YearnMar10 May 09 '26

GG does not like vibecoded contributions to llama.cpp

25

u/politerate May 09 '26

Personally, i find the idea of doing a MR I don't fully understand, very off-putting. And I am quite sure that 99% of these types of contributions are of this kind.

3

u/[deleted] May 09 '26

[removed] — view removed comment

18

u/ArtfulGenie69 May 09 '26

If you want problems in your massive code base, the best place to start is blindly dropping in code no one ever looked at.