this post was submitted on 27 Aug 2026
496 points (98.1% liked)

Technology

87624 readers
3829 users here now

This is a most excellent place for technology news and articles.


Our Rules


  1. Follow the lemmy.world rules.
  2. Only tech related news or articles.
  3. Be excellent to each other!
  4. Mod approved content bots can post up to 10 articles per day.
  5. Threads asking for personal tech support may be deleted.
  6. Politics threads may be removed.
  7. No memes allowed as posts, OK to post as comments.
  8. Only approved bots from the list below, this includes using AI responses and summaries. To ask if your bot can be added please contact a mod.
  9. Check for duplicates before posting, duplicates may be removed
  10. Accounts 7 days and younger will have their posts automatically removed.

Approved Bots


founded 3 years ago
MODERATORS
 

Damn, I hope this is not goodbye to uncensored models 😢

top 50 comments
sorted by: hot top controversial new old
[–] BeatTakeshi@lemmy.world 1 points 18 minutes ago

gets a face hugger

[–] BetterDev@programming.dev 7 points 3 hours ago

Anybody got a good list of models we should download right now before they start disappearing?

[–] Lettuceeatlettuce@lemmy.ml 2 points 2 hours ago

Pick up as many uncensored models as you can store ASAP. Those will absolutely be the first to go.

[–] flop_leash_973@lemmy.world 1 points 3 hours ago* (last edited 3 hours ago)

I would be less concerned about it being the end of uncensored models than I would be Nvidia finding ways to make sure there are no free models hosted there that work well on anything but Nvidia hardware for the foreseeable future.

[–] Gsus4@mander.xyz 27 points 18 hours ago

Ok, how come these public utility databases are owned to be sold like that eg gitgub, twitter without any regulator pushback? Oh yea...

[–] benny@reddthat.com 7 points 16 hours ago

The Spark isn't a bad computer, and they haven't completely gotten rid of local GPUs, so Nvidia does somewhat care about local llms, but power corrupts and this is just more of it.

[–] GreenKnight23@lemmy.world 5 points 16 hours ago

let the fuckening begin!

lol

[–] VirtuePacket@lemmy.zip 26 points 1 day ago
[–] melfie@lemmy.zip 62 points 1 day ago (5 children)

The core llama.cpp maintainers also work at HF and will now work for Nvidia I guess. Llama.cpp is a pretty significant part of the local LLM stack, especially since other tools like Ollama and LMStudio are just GUIs built on top.

Local LLMs have gotten to the point where they are a serious threat to Anthropic and OpenAI, and Nvidia has a lot of skin in the game. If Nvidia wanted to do some serious damage to local LLMs, they are now in a position to do so.

I’m also imagining they may try to squeeze out support for other GPU vendors. I’m using an AMD 7900 XTX to run Qwen 3.8 27B that I downloaded from HF to run on llama.cpp, which currently works like a dream. The 7900 XTX is the only sanely priced 24GB GPU left in 2026 (under $1k vs. $2k, $3k, $4k for Nvidia 24-32GB cards). Combined with OpenCode or Pi, a setup like this basically eliminates the need to use Anthropic or OpenAI products in the same way Jellyfin eliminates the need to use streaming services.

I’m sure Nvidia and their buddies don’t like one thing I’ve said in this comment and may very well be plotting to put a stop to it, so the community may need to step up our game and get our eggs out of the big tech basket.

[–] p03locke@lemmy.dbzer0.com 2 points 5 hours ago (1 children)

The core llama.cpp maintainers also work at HF and will now work for Nvidia I guess. Llama.cpp is a pretty significant part of the local LLM stack, especially since other tools like Ollama and LMStudio are just GUIs built on top.

I guess that explains why features like quantized KV caches are lagging behind. Maintainers are purposely dragging their feet.

the community may need to step up our game and get our eggs out of the big tech basket.

The community has chosen to not fight at all, which is worse. Anti-AI sentiment is at an all-time high.

Publicly. Privately, these hypocrites still whisper in ChatGPT's ear when they get lazy enough. Or use some feature in Photoshop or some other software that they didn't even understand was AI-driven.

[–] melfie@lemmy.zip 1 points 4 hours ago* (last edited 4 hours ago)

guess that explains why features like quantized KV caches are lagging behind

Yeah, I was using the TheTom fork for a while and not sure why TQ KV cache hasn’t merged yet.

Anti-AI sentiment is at an all-time high

I suppose the tech bros have understandably soured a lot of people on LLMs with all of the negative societal costs due to their greedy and irresponsible bejavior. On the other hand, the concepts of the perceptron and artificial neural networks from the 40s and 50s are finally coming to fruition and we have these things now that are legitimately artificially intelligent that we can run on our gaming PCs. They are overhyped and used in ways they make no sense, but they’re also useful as long as their limitations are kept in mind. From a technology perspective, they’re cool as hell.

[–] muusemuuse@sh.itjust.works 1 points 16 hours ago (1 children)

Oh no! -forks code- anyway…

[–] WhyJiffie@sh.itjust.works 4 points 12 hours ago

do we have experts to work on it full time, paid?

[–] boonhet@sopuli.xyz 16 points 1 day ago (4 children)

Well the good news is they can't take away from you what you already have. It being an open source project, I'm assuming if they do anything to deliberately gut AMD performance, it'll get forked.

Also

The 7900 XTX is the only sanely priced 24GB GPU left in 2026 (under $1k vs. $2k, $3k, $4k for Nvidia 24-32GB cards).

Not on sale anymore, at least not at any vendor in my country, I searched an aggregate pricing website. Amazon has a few used ones left of some models, but that's probably a 2 or 3 digit figure across SKUs. Hold on to yours with an iron grip.

What kind of tok/s are you getting with it on Qwen 3.8 27B and how's the output quality? I may consider getting one if I can find one used or import from abroad.

[–] Holytimes@sh.itjust.works 2 points 17 hours ago

I get around 40 token/s and it frequently has become reliable enough to drop sonnet for me. So take that as you will

[–] muusemuuse@sh.itjust.works 1 points 16 hours ago

Intel B50 and B60 pros are at microcenter right now perfect for this.

load more comments (2 replies)
load more comments (2 replies)
[–] echodot@feddit.uk 53 points 1 day ago* (last edited 1 day ago) (13 children)

So they are just slapping random price tags on things now. It's a database of AI questions. They have paid 13 billion dollars for a database, and everyone's acting like that's a perfectly rational sensible thing to do. No one in the financial industry has any sense anymore.

A company that's only product is something that people either don't want or actively hate has paid an eye-watering amount of money for a database to train their product on, this training will have no effect whatsoever on whether people want it.

I have this nice bridge with lots of training examples if anybody's interested, 100 trillion dollars please

[–] skisnow@lemmy.ca 1 points 8 hours ago

Yeah I don't understand these valuations at all. I wish I did so that I could get a billion dollars for a website that doesn't particularly own anything.

[–] Knock_Knock_Lemmy_In@lemmy.world 17 points 1 day ago (1 children)

for a database to train their product on

Isn't hugging face a database of already trained models?

load more comments (1 replies)
load more comments (11 replies)
[–] CosmoNova@lemmy.world 239 points 1 day ago (7 children)

Remember, the whole AI scheme exists to take the very concept of ownership from us. This takeover is hostile.

[–] cecilkorik@lemmy.ca 27 points 1 day ago (2 children)

Why do we need a centralized hub for AI models anyway? I've never figured this out. We flock to these big "friendly" fucking companies with obviously unsustainable business models trying to own and sell things that should be shared in a decentralized, democratized mesh anyway. AI models should all be distributed magnet links, not hosted files. Why do we do this to ourselves?

The way things are going at Huggingface now, I imagine we probably will have to start building the infrastructure we need to handle AI models as distributed links sooner rather than later. Let the enshittification begin, we'll move on to a different tactic for sharing AI models while cheerfully they squeeze cash out of people and businesses too lazy to adapt. Everything working as it should, I guess.

[–] p03locke@lemmy.dbzer0.com 0 points 5 hours ago

The way things are going at Huggingface now, I imagine we probably will have to start building the infrastructure we need to handle AI models as distributed links sooner rather than later.

Then "we" better get started. Mega corpos have fully embraced LLMs, to the point that they've distrusted human input too much, which is still a vital component of successful adoption. Eventually, more and more companies will wise up and figure out how to find the right balance.

Where are we at? Oh, right... we're too busy arguing about how all AI is bad and sticking our fucking heads in the sand until the Big Bad AI Problem goes away.

We have to fix our attitudes if we have any hope of surviving this mess and not ending up as a Cyberpunk-wannabe dystopia by the time 2077 hits. Use the fucking weapons given to us!

[–] avidamoeba@lemmy.ca 13 points 1 day ago* (last edited 1 day ago)

For real, this is the perfect application for torrent. If I were building a registry that had to have sustainable cost for an open source community it would use torrents as a file transfer backend. If I were building something that I was planning to sell to a monopoly firm that has monopoly money... HTTP and cloud storage all thw way! 😄

load more comments (6 replies)

It was only a matter of time

[–] DarkCloud@lemmy.world 128 points 1 day ago (4 children)

They should put up a back up torrent of all the currently available models before transferring ownership. Thus opening the doors for future websites.

NVIDIA is probably trying to shutdown the LLM-at-home market (which means you don't need huge clouds of graphics cards or even an Internet connection). Sad day for people who like to control their data.

[–] boonhet@lemmy.zip 3 points 18 hours ago* (last edited 18 hours ago)

Nvidia also sells super expensive graphics cards. They probably want you to buy a couple of 5090s or a Spark.

They know they can't shut down the entire self hosted LLM market but they can direct it towards using Nvidia GPUs by making sure llama.cpp devs don't spend much time on ROCm support, etc. And they probably know that ClosedAI and Anthropic's days on the frontier are limited.

load more comments (3 replies)
load more comments
view more: next ›