Damn, I hope this is not goodbye to uncensored models 😢
Ok, how come these public utility databases are owned to be sold like that eg gitgub, twitter without any regulator pushback? Oh yea…
Anybody got a good list of models we should download right now before they start disappearing?
They wont disappear. hugging face is two things. its a model repo and a compute provider. hugging face has paid plans to run models on their systems. nvidia will use this to run on their backend. this is how they plan to make money from hugging face. its in their best interest to keep the repo stuff as is.
The core llama.cpp maintainers also work at HF and will now work for Nvidia I guess. Llama.cpp is a pretty significant part of the local LLM stack, especially since other tools like Ollama and LMStudio are just GUIs built on top.
Local LLMs have gotten to the point where they are a serious threat to Anthropic and OpenAI, and Nvidia has a lot of skin in the game. If Nvidia wanted to do some serious damage to local LLMs, they are now in a position to do so.
I’m also imagining they may try to squeeze out support for other GPU vendors. I’m using an AMD 7900 XTX to run Qwen 3.8 27B that I downloaded from HF to run on llama.cpp, which currently works like a dream. The 7900 XTX is the only sanely priced 24GB GPU left in 2026 (under $1k vs. $2k, $3k, $4k for Nvidia 24-32GB cards). Combined with OpenCode or Pi, a setup like this basically eliminates the need to use Anthropic or OpenAI products in the same way Jellyfin eliminates the need to use streaming services.
I’m sure Nvidia and their buddies don’t like one thing I’ve said in this comment and may very well be plotting to put a stop to it, so the community may need to step up our game and get our eggs out of the big tech basket.
Well the good news is they can’t take away from you what you already have. It being an open source project, I’m assuming if they do anything to deliberately gut AMD performance, it’ll get forked.
Also
The 7900 XTX is the only sanely priced 24GB GPU left in 2026 (under $1k vs. $2k, $3k, $4k for Nvidia 24-32GB cards).
Not on sale anymore, at least not at any vendor in my country, I searched an aggregate pricing website. Amazon has a few used ones left of some models, but that’s probably a 2 or 3 digit figure across SKUs. Hold on to yours with an iron grip.
What kind of tok/s are you getting with it on Qwen 3.8 27B and how’s the output quality? I may consider getting one if I can find one used or import from abroad.
Agreed, I was referring more to future updates. Obviously we’re good with what is available now.
I can’t speak to pricing and availability outside the U.S., but it looks like the one I got went up $100:
https://www.newegg.com/asrock-radeon-rx7900xtx-24g-radeon-rx-7900-xtx-24gb-graphics-card-triple-fans/p/N82E16814930084.
I traded in my 3070 and my final price was in the 700s. Last I looked, used ones were going for $800 on eBay vs. $1200 for a used 3090.
I run 3.8 27B at q4 with q4 context up to 200k. Decode is generally in the 30s and pp starts in the 700s and drops to the 400s as context approaches 200k. I use mostly Sonnet 5 at work and I would rate this setup with the OpenCode desktop app as pretty comparable overall for coding at least. Let’s just say I have no reason to use any cloud models, not that I would do that voluntarily outside of being compelled to at work.
I get around 40 token/s and it frequently has become reliable enough to drop sonnet for me. So take that as you will
I got mine used, around 600 imperial credits. Look for ads that provide proof of working and benchmarks (like FurMark)
Intel B50 and B60 pros are at microcenter right now perfect for this.
what if the LLAMA.CPP devs working at Nvidia improves CUDA support and keeps other vendor support.
It would be great but they seem to favour server computing as that has the greatest margins. I have a feeling being the biggest company on the planet at $5tn is not enough.
Oh no! -forks code- anyway…
do we have experts to work on it full time, paid?
Remember, the whole AI scheme exists to take the very concept of ownership from us. This takeover is hostile.
Why do we need a centralized hub for AI models anyway? I’ve never figured this out. We flock to these big “friendly” fucking companies with obviously unsustainable business models trying to own and sell things that should be shared in a decentralized, democratized mesh anyway. AI models should all be distributed magnet links, not hosted files. Why do we do this to ourselves?
The way things are going at Huggingface now, I imagine we probably will have to start building the infrastructure we need to handle AI models as distributed links sooner rather than later. Let the enshittification begin, we’ll move on to a different tactic for sharing AI models while cheerfully they squeeze cash out of people and businesses too lazy to adapt. Everything working as it should, I guess.
For real, this is the perfect application for torrent. If I were building a registry that had to have sustainable cost for an open source community it would use torrents as a file transfer backend. If I were building something that I was planning to sell to a monopoly firm that has monopoly money… HTTP and cloud storage all thw way! 😄
To the one Downvoter I’m curious, why are you such a little bitch?
Probably cause it’s a dumb comment, you people see conspiracies everywhere.
That was hardly conspiracy worthy. It’s a bit hyperbolic, but it’s a fairly legit concern. Big tech is not the friend of any consumer. And big tech buying literally anything ends up a net negative for the consumer and freedom of choice. There is evidence that tech oligarchs (openAI/Sam Altman) want to gatekeep even knowledge and charge a subscription to access it.
Huggingface is now absolutely going to become pay to play. Maybe not day one, but it will happen. Just like all the other services out there that are now subscription based that were once free and offer little to justify the move. It’s also likely to drop AMD development or at least deprioritize it so hard they may as well.
There are also movements in the industry that suggest these big tech companies want to push a move to thinclients connected to cloud services instead of local machines. This idea isn’t helped by companies like Sony stripping away physical media. Nor by the RAM manufacturing cartel colluding for overpricing and abandonment of consumer markets.
Conspiracies are only conspiracies until they come to pass, and not all conspiracies are unfounded tripe like flat earthers etc.
Correct. There’s no “laughing at the idiot” button, so the downvote button has to fill in.
This comment got so much misintepreted lmao
Now there are 6 of them.
There are worse things than copyrights not mattering anymore.
Forget about rights as a whole for us plebs.
Fuck
Pick up as many uncensored models as you can store ASAP. Those will absolutely be the first to go.
Do you have a list?
We just can’t have nice things can we.
HuggingBay to the rescue 🏴☠️
You wouldn’t pirate a model
Holy shit it’s real LOL
oh that’s actually a thing, neat
I fail to understand the use of the term open source when the resource itself can be bought by oligarchs. Like organic vegetables, or clean coal.
I’ve downloaded a terabyte worth of models and didn’t pay a cent. storage and bandwidth cost money
I just put together a budget starter PC for local inference. I guess this is my cue to set up and get downloading.
It was open source till the ones who has the keys (which shouldn’t be a thing for FOSS in the first place) got an offer that made their greed take over. It’s a thing that keeps happening and it needs to be a very strong lesson that anything open source needs to be handled in a way where no one can do this.
I would be less concerned about it being the end of uncensored models than I would be Nvidia finding ways to make sure there are no free models hosted there that work well on anything but Nvidia hardware for the foreseeable future.
Get ready for the model purge
Is there something unique about hugging face that the models cannot be hosted elsewhere? I see open source as more important to Nvidia than anyone else because it sells hardware. The services’ motivations are only partially aligned with Nvidia.
I think it’s mainly just the huge amount of storage. Huggingface has been burning venture capital money for the insane among of storage space they need, afaik
I’ve never understood why they don’t seed torrents for those huge files. So much money on bandwidth that they just don’t need to spend.
Because the authors may want to take down their models or change the datasets after the fact probably.
Not that it removes from users computers, but at least the primary source can be made unavailable at will
Then remove or update the torrent link, just like any other link.
I don’t see how they could be burning considerable amounts of money on storage, as long as they have setup their own infrastructure and don’t rent it.
I don’t know how much they’re storing, I guess they’re in the petabyte territory. That’s what, a few tens of millions of dollars of up front investment?
Not free, but not such a huge amount of money.
All they do is store tho, they dont profit off that I imagine, I honestly have no clue
Probably hoping to do what GitHub did. Profit of the metrics of who is downloading what, who is developing what and having everyone uploading all their stuff to them.
Microsoft didn’t buy GitHub for the storage and bandwidth bills.
NVidia isn’t buying huggingface for the storage and bandwidth bills either
Porn
deleted by creator
The Spark isn’t a bad computer, and they haven’t completely gotten rid of local GPUs, so Nvidia does somewhat care about local llms, but power corrupts and this is just more of it.
gets a face hugger
Ed Zitron is going to have an aneurysm.
Real money or pretend AI money

That’s the neat thing, none of it’s real, and the AI investment bubble is going to make that very clear in the near future.
Or so people have been saying for the past 3 years.
Bubbles need to grow to pop






















