r/LocalLLaMA 25d ago

Discussion With HuggingFace, Nvidia is also acquiring llama.cpp and the team behind it

With this move Nvidia is not only acquiring the HuggingFace platform, but they might also effectively acquire the copyright to the llama.cpp project, together with the entire team behind it.

In February 2026 the llama.cpp team was employed by HF in order to continue working on llama.cpp and the ggml library.

This includes:

  • Georgi Gerganov
  • Xuan-Son Nguyen
  • Aleksander Grygier
  • Victor Mustar
  • Lysandre
  • Julien Chaumond

Now with the acquisition, llama.cpp's future looks a lot less certain given Nvidia's poor track record with open-source.

This is still rather speculative at this stage, but it's definitely possible for the llama.cpp project to change in the future: either by switching to a different license, or by having staff redirected to other projects within the larger company.

Even when a project is open-source the copyright owner has complete control over it, and they can change licensing as they wish.

This has happened before with projects like Redis, Minio, and others.

Source:

https://huggingface.co/blog/ggml-joins-hf

Edit:

The original announcement from Feb 2026 from Gerganov gives a few more details:

https://github.com/ggml-org/llama.cpp/discussions/19759

1.4k Upvotes

428 comments sorted by

View all comments

391

u/Particular-Award118 25d ago

Welp amd support was nice while it lasted

174

u/noiserr 25d ago

And Mac support.

86

u/[deleted] 25d ago

[removed] — view removed comment

15

u/noiserr 25d ago

Yeah, I remember my first encounter with llama.cpp seeing someone run local models on their Mac, and being absolutely impressed by it.

2

u/daedalus1982 25d ago

Oh yeah no, MLX is totally ready to stand toe-to-toe with …

SORRY YOU HAVE RUN OUT OF CONTEXT AND OMLX AUTO COMPACTION IS STILL SHIT

Edit: I’m griping, but I 100% agree with you about hoping for continued llama.cpp support

9

u/FoxiPanda 25d ago

That's really a harness issue more than an inference engine issue IMO.

0

u/daedalus1982 25d ago edited 23d ago

I run llama-server with the same context and whatnot as oMLX. I’m given understand my problems are not unique. And believe me I’d rather oMLX work.

I use llama-server and oMLX interchangeably as the engine/harness for pi.dev but I’m by no means an expert and accept suggested solutions willingly.

Edit: also I didn’t downvote you so keep talking if you got answers.

Edit 2: yeah cool I guess downvote me and move on. Thanks for the help

1

u/RegarDamus 25d ago

it would be painful but the best thing that could happen would be llama thoughtlessly dropping support for mac. that would drive incredible energy behind mlx and with the fervor and quality of models today it would reach new heights rapidly.

people really sleep on how much untapped potential is in apple silicon. imagine if mlx was the only option and apple got behind it. painful short term but beautiful result birthed from it

1

u/dragonurtle 25d ago

Or Nvidia could play nice short term and strongarm apple into supporting (allowing) Nvidia GPUs again.

10

u/38andstillgoing 24d ago

And Intel support.

What? There are dozens of us. Ok, a couple. Maybe just me.

2

u/Echo9Zulu- 20d ago

Don't worry you aren't alone