What the announcement actually says
Georgi Gerganov posted the announcement himself, as a GitHub discussion on the llama.cpp repository he founded, rather than through a press release or a filing. The text says ggml.ai, the company Gerganov set up in 2023 to fund ggml and llama.cpp, is joining Hugging Face, and that Gerganov and the rest of the team will keep spending all of their time maintaining ggml and llama.cpp. [1]
Hugging Face cross-posted the same announcement on its own blog under a near-identical title. Neither post states a purchase price, an equity stake, or any other deal term. The word used throughout is joining, not acquiring or buying. [1][2]
Atlas interpretation: The timeline lists this under acquisition, and that is a reasonable read: a company with outside investors, folded into a larger one, with its founder now working there. But the public record here is a talent and technology deal with no disclosed price, closer to what is usually called an acquihire than to a transaction with a stated valuation. Readers looking for the size of the deal will not find it in either announcement. [1][2]
The relationship predates the announcement
Gerganov's post credits two Hugging Face engineers, Xuan-Son Nguyen and Aleksander Grygier, by name for work already done inside llama.cpp before the join: core functionality contributions, a llama.cpp based inference server with its own interface, multimodal support, integration of llama.cpp into Hugging Face's Inference Endpoints, closer compatibility between the GGUF file format and the Hugging Face platform, and implementations of multiple model architectures. [1]
Atlas interpretation: That list matters more than the announcement's framing suggests. Hugging Face engineers were already committing code to llama.cpp, and llama.cpp was already wired into a Hugging Face product, before any organizational change happened. The join formalizes a dependency that already existed in both directions rather than creating one. [1]
Why this matters for how people run models locally
ggml is a tensor library built to run large models on ordinary hardware, and llama.cpp is the inference engine built on top of it that popularized running open-weight models on a laptop or phone rather than a data center GPU. The announcement calls llama.cpp the fundamental building block for local inference and Hugging Face's transformers library the fundamental building block for model definition, and describes the join as connecting those two layers. [1]
The stated technical goals are tighter compatibility between transformers model definitions and llama.cpp so a newly released model reaches local hardware faster, and better packaging so casual users can deploy local models without the current setup work. The post commits to keeping the ggml-org projects open source and community run, with the community continuing to make technical and architectural decisions on its own. [1]
Atlas interpretation: Before this, ggml.ai was a small company funding a project many people depended on and few had heard of. A widely used piece of infrastructure moving from a founder-run shop to a platform company with its own funding and commercial incentives is the kind of change that is easy to read as either good news, because the project now has a sustaining institution, or a risk, because that institution's incentives will not always match the community's. The announcement addresses the risk directly by promising the project stays open and community governed; whether that promise holds is not something a day-one announcement can prove. [1]
The institution ggml joined was itself sold seven months later
On September 3, 2026, NVIDIA announced an agreement to acquire Hugging Face, the company ggml.ai had joined a little over six months earlier. That deal has a disclosed price, roughly $12.9 billion, and is not yet closed. [3]
Atlas interpretation: This is the second ownership change the ggml and llama.cpp team has been through inside a year, and it did not choose either one. The February announcement promised long-term sustainable resources and continued community control; the September deal, if it closes, hands the parent company itself to a hardware vendor with a direct commercial stake in how local inference software behaves. Neither announcement addresses that chain, and it is worth watching rather than assuming resolved. [1][3]
Sources
- ggml.ai joins Hugging Face to ensure the long-term progress of Local AI
ggml.ai · Feb 20, 2026
- GGML and llama.cpp join HF to ensure the long-term progress of Local AI
Hugging Face · Feb 20, 2026
- NVIDIA to Acquire Hugging Face
NVIDIA · Sep 3, 2026