NVIDIA reportedly buys HuggingFace for 13 billion

https://gizmodo.com/nvidia-reportedly-stops-flirting-with-hugging-face-and-just-buys-it-2000803681

Damn, I hope this is not goodbye to uncensored models 😢

514 points · 125 comments · view on lemmy.world

125 Comments

CosmoNova@lemmy.world · 243 pts · 2d (11 replies)

Remember, the whole AI scheme exists to take the very concept of ownership from us. This takeover is hostile.

cecilkorik@lemmy.ca · 28 pts · 2d (3 replies)

Why do we need a centralized hub for AI models anyway? I've never figured this out. We flock to these big "friendly" fucking companies with obviously unsustainable business models trying to own and sell things that should be shared in a decentralized, democratized mesh anyway. AI models should all be distributed magnet links, not hosted files. Why do we do this to ourselves?

The way things are going at Huggingface now, I imagine we probably will have to start building the infrastructure we need to handle AI models as distributed links sooner rather than later. Let the enshittification begin, we'll move on to a different tactic for sharing AI models while cheerfully they squeeze cash out of people and businesses too lazy to adapt. Everything working as it should, I guess.

avidamoeba@lemmy.ca · 15 pts · 2d

For real, this is the perfect application for torrent. If I were building a registry that had to have sustainable cost for an open source community it would use torrents as a file transfer backend. If I were building something that I was planning to sell to a monopoly firm that has monopoly money... HTTP and cloud storage all thw way! 😄

p03locke@lemmy.dbzer0.com · 1 pts · 1d (1 reply)

The way things are going at Huggingface now, I imagine we probably will have to start building the infrastructure we need to handle AI models as distributed links sooner rather than later.

Then "we" better get started. Mega corpos have fully embraced LLMs, to the point that they've distrusted human input too much, which is still a vital component of successful adoption. Eventually, more and more companies will wise up and figure out how to find the right balance.

Where are we at? Oh, right... we're too busy arguing about how all AI is bad and sticking our fucking heads in the sand until the Big Bad AI Problem goes away.

We have to fix our attitudes if we have any hope of surviving this mess and not ending up as a Cyberpunk-wannabe dystopia by the time 2077 hits. Use the fucking weapons given to us!

cecilkorik@lemmy.ca · 1 pts · 18h

Sounds about right, my friend. I'll meet you in the Badlands with the other nomads after the blackwall goes up!

woelkchen@lemmy.world · 13 pts · 2d (1 reply)

There are worse things than copyrights not mattering anymore.

CosmoNova@lemmy.world · 19 pts · 2d

Forget about rights as a whole for us plebs.

a_non_monotonic_function@lemmy.world · -24 pts · 2d (4 replies)

To the one Downvoter I'm curious, why are you such a little bitch?

Axolotl_cpp@feddit.it · 3 pts · 2d (1 reply)

This comment got so much misintepreted lmao

a_non_monotonic_function@lemmy.world · -1 pts · 2d

Now there are 6 of them.

SorryQuick@lemmy.ca · -3 pts · 1d (1 reply)

Probably cause it’s a dumb comment, you people see conspiracies everywhere.

Epp@lemmus.org · 1 pts · 17h

Correct. There's no "laughing at the idiot" button, so the downvote button has to fill in.

DarkCloud@lemmy.world · 131 pts · 2d (4 replies)

They should put up a back up torrent of all the currently available models before transferring ownership. Thus opening the doors for future websites.

NVIDIA is probably trying to shutdown the LLM-at-home market (which means you don't need huge clouds of graphics cards or even an Internet connection). Sad day for people who like to control their data.

ThePowerOfGeek@lemmy.world · 43 pts · 2d (1 reply)

I heard somewhere they are buying it because their own development platform sucks. So they are going to replace it with hugging face.

Which doesn't preclude them from locking all the good stuff behind a subscription, of course.

dil@lemmy.zip · 5 pts · 2d

Maybe a dev fee to post

byte_0verflow@lemmy.ml · 10 pts · 2d

Before any of these news I would use modelscope occasionally, it is the same thing as huggingface but Chinese and with a massive collection of mcp servers. Although they do not have the same amount of merged and fine tuned models, they are still a pretty good service

boonhet@lemmy.zip · 3 pts · 1d

Nvidia also sells super expensive graphics cards. They probably want you to buy a couple of 5090s or a Spark.

They know they can't shut down the entire self hosted LLM market but they can direct it towards using Nvidia GPUs by making sure llama.cpp devs don't spend much time on ROCm support, etc. And they probably know that ClosedAI and Anthropic's days on the frontier are limited.

TropicalDingdong@lemmy.world · 95 pts · 2d (6 replies)

We just can't have nice things can we.

obsidian@discuss.online · 98 pts · 2d (4 replies)

HuggingBay to the rescue 🏴‍☠️

Semi_Hemi_Demigod@lemmy.world · 49 pts · 2d (1 reply)

You wouldn’t pirate a model

W98BSoD@lemmy.dbzer0.com · 33 pts · 2d

eager_eagle@lemmy.world · 18 pts · 2d

oh that's actually a thing, neat

87Six@lemmy.zip · 1 pts · 19h

Holy shit it's real LOL

p03locke@lemmy.dbzer0.com · 1 pts · 1d

We are in a war with the rich. Always have been. We will never have nice things unless you take it from them.

frustrated_phagocytosis@fedia.io · 80 pts · 2d (3 replies)

I fail to understand the use of the term open source when the resource itself can be bought by oligarchs. Like organic vegetables, or clean coal.

darkkite@lemmy.ml · 13 pts · 2d (1 reply)

I've downloaded a terabyte worth of models and didn't pay a cent. storage and bandwidth cost money

mrunicornman@lemmy.world · 3 pts · 2d

I just put together a budget starter PC for local inference. I guess this is my cue to set up and get downloading.

M137@lemmy.today · 1 pts · 1d

It was open source till the ones who has the keys (which shouldn't be a thing for FOSS in the first place) got an offer that made their greed take over. It's a thing that keeps happening and it needs to be a very strong lesson that anything open source needs to be handled in a way where no one can do this.

civ@lemmy.civl.cc · 65 pts · 2d (10 replies)

Get ready for the model purge

olympicyes@lemmy.world · 22 pts · 2d (9 replies)

Is there something unique about hugging face that the models cannot be hosted elsewhere? I see open source as more important to Nvidia than anyone else because it sells hardware. The services’ motivations are only partially aligned with Nvidia.

civ@lemmy.civl.cc · 30 pts · 2d (6 replies)

I think it's mainly just the huge amount of storage. Huggingface has been burning venture capital money for the insane among of storage space they need, afaik

Serinus@lemmy.world · 13 pts · 2d (2 replies)

I've never understood why they don't seed torrents for those huge files. So much money on bandwidth that they just don't need to spend.

boonhet@lemmy.zip · 1 pts · 1d (1 reply)

Because the authors may want to take down their models or change the datasets after the fact probably.

Not that it removes from users computers, but at least the primary source can be made unavailable at will

Serinus@lemmy.world · 4 pts · 23h

Then remove or update the torrent link, just like any other link.

ranzispa@mander.xyz · 3 pts · 2d (2 replies)

I don't see how they could be burning considerable amounts of money on storage, as long as they have setup their own infrastructure and don't rent it.

I don't know how much they're storing, I guess they're in the petabyte territory. That's what, a few tens of millions of dollars of up front investment?

Not free, but not such a huge amount of money.

dil@lemmy.zip · 2 pts · 2d (1 reply)

All they do is store tho, they dont profit off that I imagine, I honestly have no clue

lefaucet@slrpnk.net · 4 pts · 2d

Probably hoping to do what GitHub did. Profit of the metrics of who is downloading what, who is developing what and having everyone uploading all their stuff to them.

Microsoft didn't buy GitHub for the storage and bandwidth bills.

NVidia isn't buying huggingface for the storage and bandwidth bills either

socsa@piefed.social · 2 pts · 2d

Porn

KRAW@linux.community · 1 pts · 2d
[ removed ]
melfie@lemmy.zip · 63 pts · 2d (11 replies)

The core llama.cpp maintainers also work at HF and will now work for Nvidia I guess. Llama.cpp is a pretty significant part of the local LLM stack, especially since other tools like Ollama and LMStudio are just GUIs built on top.

Local LLMs have gotten to the point where they are a serious threat to Anthropic and OpenAI, and Nvidia has a lot of skin in the game. If Nvidia wanted to do some serious damage to local LLMs, they are now in a position to do so.

I’m also imagining they may try to squeeze out support for other GPU vendors. I’m using an AMD 7900 XTX to run Qwen 3.8 27B that I downloaded from HF to run on llama.cpp, which currently works like a dream. The 7900 XTX is the only sanely priced 24GB GPU left in 2026 (under $1k vs. $2k, $3k, $4k for Nvidia 24-32GB cards). Combined with OpenCode or Pi, a setup like this basically eliminates the need to use Anthropic or OpenAI products in the same way Jellyfin eliminates the need to use streaming services.

I’m sure Nvidia and their buddies don’t like one thing I’ve said in this comment and may very well be plotting to put a stop to it, so the community may need to step up our game and get our eggs out of the big tech basket.

boonhet@sopuli.xyz · 16 pts · 2d (4 replies)

Well the good news is they can't take away from you what you already have. It being an open source project, I'm assuming if they do anything to deliberately gut AMD performance, it'll get forked.

Also

The 7900 XTX is the only sanely priced 24GB GPU left in 2026 (under $1k vs. $2k, $3k, $4k for Nvidia 24-32GB cards).

Not on sale anymore, at least not at any vendor in my country, I searched an aggregate pricing website. Amazon has a few used ones left of some models, but that's probably a 2 or 3 digit figure across SKUs. Hold on to yours with an iron grip.

What kind of tok/s are you getting with it on Qwen 3.8 27B and how's the output quality? I may consider getting one if I can find one used or import from abroad.

melfie@lemmy.zip · 4 pts · 2d

Agreed, I was referring more to future updates. Obviously we’re good with what is available now.

I can’t speak to pricing and availability outside the U.S., but it looks like the one I got went up $100:

https://www.newegg.com/asrock-radeon-rx7900xtx-24g-radeon-rx-7900-xtx-24gb-graphics-card-triple-fans/p/N82E16814930084. 

I traded in my 3070 and my final price was in the 700s. Last I looked, used ones were going for $800 on eBay vs. $1200 for a used 3090.

I run 3.8 27B at q4 with q4 context up to 200k. Decode is generally in the 30s and pp starts in the 700s and drops to the 400s as context approaches 200k. I use mostly Sonnet 5 at work and I would rate this setup with the OpenCode desktop app as pretty comparable overall for coding at least. Let’s just say I have no reason to use any cloud models, not that I would do that voluntarily outside of being compelled to at work.

adhdsergio@lemmy.world · 3 pts · 2d

I got mine used, around 600 imperial credits. Look for ads that provide proof of working and benchmarks (like FurMark)

Holytimes@sh.itjust.works · 2 pts · 1d

I get around 40 token/s and it frequently has become reliable enough to drop sonnet for me. So take that as you will

muusemuuse@sh.itjust.works · 1 pts · 1d

Intel B50 and B60 pros are at microcenter right now perfect for this.

Mwa@thelemmy.club · 4 pts · 2d (1 reply)

what if the LLAMA.CPP devs working at Nvidia improves CUDA support and keeps other vendor support.

adhdsergio@lemmy.world · 3 pts · 2d

It would be great but they seem to favour server computing as that has the greatest margins. I have a feeling being the biggest company on the planet at $5tn is not enough.

p03locke@lemmy.dbzer0.com · 3 pts · 1d (1 reply)

The core llama.cpp maintainers also work at HF and will now work for Nvidia I guess. Llama.cpp is a pretty significant part of the local LLM stack, especially since other tools like Ollama and LMStudio are just GUIs built on top.

I guess that explains why features like quantized KV caches are lagging behind. Maintainers are purposely dragging their feet.

the community may need to step up our game and get our eggs out of the big tech basket.

The community has chosen to not fight at all, which is worse. Anti-AI sentiment is at an all-time high.

Publicly. Privately, these hypocrites still whisper in ChatGPT's ear when they get lazy enough. Or use some feature in Photoshop or some other software that they didn't even understand was AI-driven.

melfie@lemmy.zip · 1 pts · 1d
[ removed ]
muusemuuse@sh.itjust.works · 1 pts · 1d (1 reply)

Oh no! -forks code- anyway…

WhyJiffie@sh.itjust.works · 4 pts · 1d

do we have experts to work on it full time, paid?

echodot@feddit.uk · 57 pts · 2d (21 replies)

So they are just slapping random price tags on things now. It's a database of AI questions. They have paid 13 billion dollars for a database, and everyone's acting like that's a perfectly rational sensible thing to do. No one in the financial industry has any sense anymore.

A company that's only product is something that people either don't want or actively hate has paid an eye-watering amount of money for a database to train their product on, this training will have no effect whatsoever on whether people want it.

I have this nice bridge with lots of training examples if anybody's interested, 100 trillion dollars please

Knock_Knock_Lemmy_In@lemmy.world · 17 pts · 2d (1 reply)

for a database to train their product on

Isn't hugging face a database of already trained models?

sekki@lemmy.world · 9 pts · 2d

Partially. But they also host datasets.

skisnow@lemmy.ca · 2 pts · 1d

Yeah I don't understand these valuations at all. I wish I did so that I could get a billion dollars for a website that doesn't particularly own anything.

General_Effort@lemmy.world · 1 pts · 18h

Is this some attempt at poisoning AI?

1985MustangCobra@lemmy.ca · -2 pts · 1d (16 replies)

not everyone has a hate boner for AI

echodot@feddit.uk · 2 pts · 1d (15 replies)

It's an expensive toy that might become a viable product in a decade that doesn't mean it's a valuable company today. Irresponsible spending like this is exactly what led to the dot com bubble.

1985MustangCobra@lemmy.ca · -1 pts · 1d (14 replies)

its a useful tool for me doing research.

Alcoholicorn@mander.xyz · 5 pts · 1d (7 replies)

What kind of research do llms actually help with?

In my experience, if the answer can't be found in the top google results, the AI simply makes shit up.

echodot@feddit.uk · 4 pts · 1d (1 reply)

The fact that it lies is so irritating, if it at least admitted that it didn't know the answer it would at least be a useful search engine. The fact that it makes stuff up though means I always have to scroll past the AI summary, that I didn't ask for, in order to actually go to the website to check the results myself, because I can't trust it.

Alcoholicorn@mander.xyz · 3 pts · 1d

The worst part is that mostly lies for things that aren't immediately obvious, so you ask it "what is the color of the sun" and it appears to work just fine, but the moment you use it for any real research, it will just fabricate things wholecloth.

1985MustangCobra@lemmy.ca · 0 pts · 1d (4 replies)

thats completely false when you only ask the llm a very basic question like "what color is the sun?" you need to talk to the model like its smarter than a google search engine. "what color is the sun? pull the data from scientific articles and other sites that study the sun"

if you google the first question, you will have list of different pages, where if you use AI to search on your query, it will model an answer from thoese sources, with said sources linked, similar to a Wikipedia page.

Alcoholicorn@mander.xyz · 4 pts · 1d (3 replies)

There are thousands of papers on the color of the sun, its trivial to google a paper, if you need an LLM to do this, that's on you. But if you ask it valve clearances for an indonesian motorbike in english, it will make it up, even though the service manual is available in Indonesian. If you ask it how to synthesize a chemical that nobody bothers to write about because its uninteresting or impossible, it will straight up lie.

1985MustangCobra@lemmy.ca · -3 pts · 1d

i think that learning to prompt AI is somthing people lack as a skill and then call it bad and a liar like its some intelligence when it isn't

echodot@feddit.uk · 1 pts · 1d (5 replies)

Yeah because it was impossible to Google things in the past. Every time anybody comes up with a use for AI it's basically just automating something that isn't even that hard for you to do. If you want to automate simple tasks that's absolutely fine but it's not the second coming as people keep insisting.

Come back to me when an AI can actually add value to something, when it can do something that is not just automating a simple task, but it's capable of doing things that humans are not. I keep being told that we're only a few years away from the technological singularity, so wake me up when we actually get there.

ThirdConsul@lemmy.zip · 1 pts · 19h (1 reply)

when it can do something that is not just automating a simple task, but it’s capable of doing things that humans are no

No, credit when credit's due, LLMs have arrived (by themselves) or helped to arrive at some mathematical conclusions, didn't they?

A few of Erdős, some others.

Yes, it's, a very small subset (duh), out of god knows how many they tried, and achieving the solution is made via technique called "infinite monkeys typing out Hamlet", but credit when credit's due (unlike LLMs that don't properly attribute parts of the proofs).

echodot@feddit.uk · 1 pts · 16h

Yeah I knew you'd say something like that. Mean while the rest of the human race is busy getting on with existing and not worrying about hyper specific mathematical constructs that have no bearing on actual reality. Let me know when you come up with some kind of actual concrete used for the technology rather than some toy situation.

Epp@lemmus.org · 0 pts · 17h (2 replies)

It saves immeasurable time. Your time may not be valuable, but mine is.

echodot@feddit.uk · 1 pts · 16h (1 reply)

Yeah and another way to say that is my job is complicated enough they cannot be automated via lua script with ambition.

Rather than being an ass wipe why don't you tell us what you are immensely important and easily automatable job is. And then I guess you can go find a new career, or something.

Epp@lemmus.org · 1 pts · 15h

I've posted examples in the past, enema bag. No new career necessary. My current career, attained with an MSc in Machine Learning, is serving me just fine.

KiwiTB@lemmy.world · 40 pts · 2d (3 replies)

Real money or pretend AI money

db2@lemmy.world · 28 pts · 2d

naught101@lemmy.world · 8 pts · 2d (1 reply)

That's the neat thing, none of it's real, and the AI investment bubble is going to make that very clear in the near future.

SorryQuick@lemmy.ca · 3 pts · 1d

Or so people have been saying for the past 3 years.

aesthelete@lemmy.world · 33 pts · 2d

Ed Zitron is going to have an aneurysm.

01189998819991197253@infosec.pub · 31 pts · 2d

Gsus4@mander.xyz · 27 pts · 2d (2 replies)

Ahhhh, this is what they meant by "the circular economy"

frunch@lemmy.world · 13 pts · 2d

Circular like this

WhatAmLemmy@lemmy.world · 2 pts · 2d

It's more of a reverse funnel system. A pyramid scheme, if you will.

VirtuePacket@lemmy.zip · 26 pts · 1d

Fuck

Gsus4@mander.xyz · 26 pts · 1d

Ok, how come these public utility databases are owned to be sold like that eg gitgub, twitter without any regulator pushback? Oh yea...

negativenull@piefed.world · 24 pts · 2d (3 replies)

My question is what their motivation is.

  • Do they want to close down open models so more is directed at the frontier/closed companies (OpenAI/Anthropic)?
  • Or are they trying to diversify and push open models as a safety net for when the bubble bursts
DudeImMacGyver@kbin.earth · 40 pts · 2d

The latter sounds like the smarter move, so it's probably the former.

yeh74fjic8e5we@lemmy.world · 3 pts · 1d

To stop a competitor from getting the same kind of control. Particularly if its one of the foreign (to US) ones that's less amenable to whatever US policy is on any given day.

skisnow@lemmy.ca · 2 pts · 1d

One of the many reasons is likely to make sure models stay specifically optimized for NVidia hardware, by floating them to the top of the search results and making sure that they're the only ones that get special prices on the tools and hosting. It's such a fast-moving industry that they know they won't always have the fastest chips indefinitely.

einlander@lemmy.world · 19 pts · 2d

And this will be why the Chinese models will win.

homesweethomeMrL@lemmy.world · 19 pts · 2d (1 reply)

Oh no we have to get decent local models from somewhere else now.

leanleft@lemmy.ml · 13 pts · 2d

this site seems to be better than others ive seen : https://llama.garden/
complaints - the torrents are packs of all different quants .. or some are just safetensor files.

humanspiral@lemmy.ca · 18 pts · 2d (7 replies)

I know huggingface. I don't know of anything they sell, though.

REDACTED@infosec.pub · 8 pts · 2d (6 replies)

Neither does github

FlexibleToast@lemmy.world · 4 pts · 1d (2 replies)

I've worked at plenty of places with github enterprise subscriptions. They definitely sell things, you're just not the consumer they care about.

skisnow@lemmy.ca · 1 pts · 1d (1 reply)

seconded, they charge $21/month/seat if you're using the pro features; that's $50k a year even for a smallish team of 20, just to use their website. Once you get to a bigger company with 800 seats, that's $1M over 5 years, for something that was mostly built on free OSS software to begin with.

Not to mention the biggest bait-and-switch in history that they pulled off on Github Copilot subscriptions, that I'm guessing a lot of big orgs won't have cancelled yet. They really are laughing all the way to the bank.

BlaestEgnen@feddit.dk · 1 pts · 1d

20*21*12 is more than 50k?

boonhet@lemmy.zip · 1 pts · 1d (1 reply)

? Github can cost tons of money if you abuse runners

GreenKnight23@lemmy.world · 1 pts · 1d

isn't that just github?

GreenKnight23@lemmy.world · 1 pts · 1d

github sold premium services.

then Microsoft bought it.

cyberpunk007@lemmy.ca · 17 pts · 2d (3 replies)

Buys what?

Knock_Knock_Lemmy_In@lemmy.world · 12 pts · 2d (2 replies)

Buys the ability to control open source model distribution.

cyberpunk007@lemmy.ca · 1 pts · 2d (1 reply)

Oh no.

boonhet@lemmy.zip · 1 pts · 1d

The good news is that there's nothing stopping us from starting a competitor. Nvidia probably wants to use it to push their graphics cards.

ramenshaman@lemmy.world · 14 pts · 2d (5 replies)

ELI5 huggingface?

NGC2346@sh.itjust.works · 25 pts · 2d (1 reply)

Its a big box of toys but replace the toys with downloadable AI models and the box is a website

ramenshaman@lemmy.world · 1 pts · 16h

Wow that was an actual ELI5, thanks!

dil@lemmy.zip · 10 pts · 2d (1 reply)

ngl the first image on their website is pretty much all the explanation you need

GreenKnight23@lemmy.world · 1 pts · 1d

I am a LLM, what is this?

dil@lemmy.zip · 8 pts · 2d

hosts all the ai models and is the primary resource for downloading them for pretty much all kinds, like depthmaps, masking too, also llms, inage generators, etc.

sns@lemmy.dbzer0.com · 14 pts · 2d

The Ouroboros must be fed!

thedeadwalking4242@lemmy.world · 11 pts · 2d

It was only a matter of time

Sims@lemmy.ml · 10 pts · 2d
Hackworth@piefed.ca · 8 pts · 2d

One of the few AI orgs of note outside of China and the US. What's left? Black Forest?

BetterDev@programming.dev · 7 pts · 1d (1 reply)

Anybody got a good list of models we should download right now before they start disappearing?

terabyterex@lemmy.world · 2 pts · 18h

They wont disappear. hugging face is two things. its a model repo and a compute provider. hugging face has paid plans to run models on their systems. nvidia will use this to run on their backend. this is how they plan to make money from hugging face. its in their best interest to keep the repo stuff as is.

Sibshops@feddit.cl · 7 pts · 2d

This doesn't sound like it will help the trout population.

benny@reddthat.com · 7 pts · 1d

The Spark isn't a bad computer, and they haven't completely gotten rid of local GPUs, so Nvidia does somewhat care about local llms, but power corrupts and this is just more of it.

Buffalox@lemmy.world · 5 pts · 2d

AFAIK that's about 2 weeks revenue. IDK what their profit margin is, but I bet it's high.

GreenKnight23@lemmy.world · 5 pts · 1d

let the fuckening begin!

lol

Mwa@thelemmy.club · 5 pts · 2d (3 replies)

Please tell me this is a rumor

adhdsergio@lemmy.world · 4 pts · 2d (2 replies)

It's basically a done deal

Mwa@thelemmy.club · 8 pts · 2d (1 reply)

Fuck

adhdsergio@lemmy.world · 2 pts · 1d

Aye

flop_leash_973@lemmy.world · 2 pts · 23h

I would be less concerned about it being the end of uncensored models than I would be Nvidia finding ways to make sure there are no free models hosted there that work well on anything but Nvidia hardware for the foreseeable future.

BeatTakeshi@lemmy.world · 2 pts · 20h

gets a face hugger

Lettuceeatlettuce@lemmy.ml · 2 pts · 23h (1 reply)

Pick up as many uncensored models as you can store ASAP. Those will absolutely be the first to go.

MalMen@sh.itjust.works · 1 pts · 17h

Do you have a list?

rotkehle@feddit.org · 1 pts · 2d
[ removed ]
obsidian@discuss.online · 0 pts · 2d
[ removed ]