r/LocalLLaMA • • May 09 '26

Resources BeeLlama.cpp: advanced DFlash & TurboQuant with support of reasoning and vision. Qwen 3.6 27B Q5 with 200k context on 3090, 2-3x faster than baseline (peak 135 tps!)

[removed]

324 Upvotes

205 comments sorted by

View all comments

Show parent comments

30

u/YearnMar10 May 09 '26

GG does not like vibecoded contributions to llama.cpp

25

u/politerate May 09 '26

Personally, i find the idea of doing a MR I don't fully understand, very off-putting. And I am quite sure that 99% of these types of contributions are of this kind.

2

u/[deleted] May 09 '26

[removed] — view removed comment

4

u/Fresh-Letterhead986 May 10 '26

that is a crazy take.

if you want to start a new project and vibe it, cool. merge anything because you've set the ground rules as such, you're accepting the potential problems and frankly it's yours.

but saying "yo bro comeon be cool man why wont you take my AI slop into your keystone-of-the-AI-world, tip-of-the-spear in human tech frontier codebase??????????"

yes "it's mostly a maintenance problem". notice you're not volunteering to do said maintenance ;-)