Back to latest

The company that owns the AI compute stack is now moving on the layer where the industry …

Nvidia is in advanced talks to acquire Hugging Face, according to a report circulated Tuesday — a deal that, if it closes, would put the CUDA vendor in control of the hub …

Top Story

The company that owns the AI compute stack is now moving on the layer where the industry actually gets its models. Nvidia is in advanced talks to acquire Hugging Face, according to a report circulated Tuesday — a deal that, if it closes, would put the CUDA vendor in control of the hub where most engineers find, download, fine-tune, and serve open-weight models. Neither company has confirmed the talks, and the only available “report” is a bare link with no sourcing, so treat the figure and the timeline as unconfirmed. What is not in doubt is the shape of the deal and why it matters: Nvidia already sells the chips, the software layer, and increasingly the inference servers that run frontier AI. Hugging Face is the distribution channel through which a large share of that work starts.

Hugging Face is the GitHub of models, but the analogy understates it. The Hub is not just a registry; it hosts model weights, datasets, Spaces demo apps, and a widely used inference API that developers point at without standing up their own hardware. A meaningful portion of the open-model traffic that reaches Nvidia silicon — and reaches competitors’ silicon too — first passes through Hugging Face. That position is valuable precisely because it is neutral. Meta ships Llama there. Microsoft, Google, and Mistral distribute on it. Startups treat it as the default first stop before they build anything, and its Inference Providers program is quietly becoming a toll booth for running other people’s models.

That neutrality is what Nvidia would be buying, and what it would put at risk. The company already enjoys a structural position no other hardware vendor has matched: CUDA locks developers into its platform, and it owns the supply chain for training and inference accelerators. Owning the distribution layer on top of that is the AI-era equivalent of the platform company that buys the app store. The obvious hazard is that Hugging Face’s biggest customers are also Nvidia’s biggest rivals’ open-model programs. Meta has been cultivating Hugging Face as the canonical home for Llama; if that home is owned by the chip company whose accelerators Meta is trying to disintermediate, the arrangement starts to look fragile.

The deal lands on a day when the rest of the stack is consolidating under fewer hands in ways builders should read as one pattern rather than separate headlines. Google shipped Gemini 3.8 Flash, tightening its grip on cheap frontier inference. The administration is backing OpenAI in its copyright fight with the New York Times, tilting the ground rules for training data in the direction of the incumbents who can afford to litigate. And Mistral flipped its default so that user input trains its models unless a customer is on the enterprise tier — a quiet redefinition of what “sharing” your prompt data means. Model distribution, data defaults, and copyright rules are all moving at once, and each move concentrates more leverage in the vendors who own the rails rather than the developers who ride them.

Nvidia buying Hugging Face is the most structural of these because it fuses the physical and the social layers of the stack. There is a real tension in the plan: Hugging Face’s community value depends on being a place where everyone ships, and a single owner with Nvidia’s commercial interests will test how much of that goodwill survives an acquisition. The regulators who have already scrutinized Nvidia’s dominance in datacenter accelerators will have something to say about whether the maker of the chips should also own the default clearinghouse for the models that run on them.

Watch Meta. Llama is the Hub’s biggest open-weight draw, and Meta has treated Hugging Face as the de facto distribution arm for its most important research release. If Nvidia takes ownership and Meta starts routing Llama’s distribution through its own channel — or quietly downgrades the Hub’s role — the acquisition’s strategic value drops as fast as its optics deteriorate. The company with the largest open models is the one whose next distribution decision will tell you whether a Nvidia-owned Hugging Face is a moat or a poisoned well. Nvidia In Advanced Talks for Hugging Face

Also Today

Gemini 3.8 Flash and 3.8 Flash Cyber · Source Google shipped its third Flash model in six weeks, putting Gemini 3.8 Flash — reasoning and coding gains over 3.7 at the same $0.75/$3.75 per-million-token price — alongside a gated Cyber variant that, per Google, finds 2.6x more correct Chrome patches than larger commercial rivals and hit a 47.2% pass@1 on CWE-Bench against a frontier model’s 47.8%. The Cyber build is locked behind the new Fairwind Program for governments and critical-infrastructure defenders, a deliberate gating of offensive-capable weights. But the cadence matters as much as the numbers: a price-locked Flash every few weeks is Google using cost discipline to own the mid-tier.

The Trump administration is supporting OpenAI in the NYT copyright lawsuit · Source The Department of Justice formally took OpenAI’s side in the New York Times suit, filing a letter that calls LLM training on copyrighted articles “extraordinarily transformative” and warns that restraining it “would thwart such creative and scientific progress.” It lands when the copyright picture is otherwise split — Anthropic just settled for a record $1.5 billion while Meta won Kadrey largely on an evidence technicality, and Sony and Warner added fresh suits last week. The letter is advisory, not binding on Judge Stein, but a DOJ fair-use position is the administration spelling out which way it wants AI law to lean.

Mistral now trains on user input by default, except on enterprise tier · Source Mistral now trains on user input and output by default everywhere except the enterprise tier, where an admin must actively opt in. On consumer Vibe, conversations, uploaded documents and outputs feed model training unless a user disables a toggle in the admin panel, with separate settings for Vibe and the API that must each be configured independently. It is an explicit reversal of the opt-in norm that shifts the burden to the user to find two different panels. The engineering floor will be the one that notices, because it reads the defaults first.

Three sites made 215,128 “best software” pages for AI. Perplexity cites them · Source An audit of what Perplexity actually grounds on found that of 7,534 citations across 380 buyer-intent queries, 59.8% point to domains ranked worse than #100,000 and 23.4% to sites outside the top million entirely. Three apparently common-owned brands that title their homepages “Facts & Grounding Page” in HTML have published 215,128 generated best-software guides between them, and guideflow.com — a demo vendor in none of the categories asked — was the third-most-cited source, ahead of Gartner. The retrieval layer is not stumbling onto these pages; the pages are written for it.

We could save petabytes of cache storage with Zstandard and Pingora · Source Cloudflare prototyped a cache layer that zstd-compresses eligible text assets before writing them to disk, shrinking them to about a third of on-disk size at level 3 while holding extra CPU to a few percent. Text is 22.3% of bytes but 67.3% of requests, and 71% of it arrives uncompressed, so the win concentrates where reuse is high and the 4 KiB floor drops only ~1% of eligible bytes. The economics are the point: encode once per fill, then save on storage and inter-datacenter transfer on every serve of a hot asset.

In Brief

  • The FBI said it is investigating a dark-web identity service advertising scans of more than 153 million U.S. and Canadian drivers licenses. (Source)
  • Delivery Hero’s board voted to accept Uber’s roughly $15 billion takeover offer, moving the deal a step closer to closing. (Source)
  • Google opened the Fairwind Program, a limited-access channel giving governments and critical-infrastructure operators first crack at Gemini 3.8 Flash Cyber. (Source)
  • 1Password is facing customer and employee backlash over a $300,000 pledge to a Linux distribution created by David Heinemeier Hansson, who has published racist and anti-immigrant commentary. (Source)
  • The FCC proposed a public “robocall mitigation scorecard” grading phone companies on how well they block spam calls while avoiding false positives on legitimate ones. (Source)
  • Dell shares rose after it raised its sales outlook for a fifth straight quarter on surging demand for AI and traditional servers. (Source)
  • A paper argues modern neural networks exhibit an emergent symbolic structure over structured combinations of concepts, despite being trained with no explicit symbolic machinery. (Source)
  • Anthropic’s Claude Fable 5.1 now produces explorable, browser-native reconstructions of real places that are researched, modeled, and quality-checked end to end by the model. (Source)
  • A note on AI coding economics argues real efficiency is cutting the token count of individual agent runs rather than chasing marginal task-quality gains. (Source)
  • A short technical note connects OpenAI Astra to recurrent depth, looped transformers, Nanbeige 4.2, and the recent Mixture-of-Recursions paper. (Source)
  • A reply to Matklad’s “Memory Safety’s Hardest Problem” makes the case for static allocation and constant work as the way out. (Source)
  • Nuxt 4.5, billed as its largest release in a while, adds experimental SSR streaming, Vite 8, and an Rsbuild-based Rspack builder. (Source)

One Line

The Administration is siding with a handful of trillion-dollar AI companies at the expense of the countless American creators whose work they stole.

— New York Times spokesperson Graham James, quoted by WIRED on the DOJ’s fair-use letter in the OpenAI copyright suit.

Search MacWorks

Enter at least two characters.