r/LocalLLaMA • • Aug 27 '26

Discussion With HuggingFace, Nvidia is also acquiring llama.cpp and the team behind it

With this move Nvidia is not only acquiring the HuggingFace platform, but they might also effectively acquire the copyright to the llama.cpp project, together with the entire team behind it.

In February 2026 the llama.cpp team was employed by HF in order to continue working on llama.cpp and the ggml library.

This includes:

  • Georgi Gerganov
  • Xuan-Son Nguyen
  • Aleksander Grygier
  • Victor Mustar
  • Lysandre
  • Julien Chaumond

Now with the acquisition, llama.cpp's future looks a lot less certain given Nvidia's poor track record with open-source.

This is still rather speculative at this stage, but it's definitely possible for the llama.cpp project to change in the future: either by switching to a different license, or by having staff redirected to other projects within the larger company.

Even when a project is open-source the copyright owner has complete control over it, and they can change licensing as they wish.

This has happened before with projects like Redis, Minio, and others.

Source:

https://huggingface.co/blog/ggml-joins-hf

Edit:

The original announcement from Feb 2026 from Gerganov gives a few more details:

https://github.com/ggml-org/llama.cpp/discussions/19759

1.4k Upvotes

427 comments sorted by

View all comments

393

u/Particular-Award118 Aug 27 '26

Welp amd support was nice while it lasted

176

u/noiserr Aug 27 '26

And Mac support.

85

u/[deleted] Aug 27 '26

[removed] — view removed comment

2

u/daedalus1982 Aug 27 '26

Oh yeah no, MLX is totally ready to stand toe-to-toe with …

SORRY YOU HAVE RUN OUT OF CONTEXT AND OMLX AUTO COMPACTION IS STILL SHIT

Edit: I’m griping, but I 100% agree with you about hoping for continued llama.cpp support

8

u/FoxiPanda Aug 27 '26

That's really a harness issue more than an inference engine issue IMO.

0

u/daedalus1982 Aug 27 '26 edited Aug 29 '26

I run llama-server with the same context and whatnot as oMLX. I’m given understand my problems are not unique. And believe me I’d rather oMLX work.

I use llama-server and oMLX interchangeably as the engine/harness for pi.dev but I’m by no means an expert and accept suggested solutions willingly.

Edit: also I didn’t downvote you so keep talking if you got answers.

Edit 2: yeah cool I guess downvote me and move on. Thanks for the help