this post was submitted on 09 Apr 2026
371 points (98.4% liked)

Science Memes

19839 readers
2561 users here now

Welcome to c/science_memes @ Mander.xyz!

A place for majestic STEMLORD peacocking, as well as memes about the realities of working in a lab.



Rules

  1. Don't throw mud. Behave like an intellectual and remember the human.
  2. Keep it rooted (on topic).
  3. No spam.
  4. Infographics welcome, get schooled.

This is a science community. We use the Dawkins definition of meme.



Research Committee

Other Mander Communities

Science and Research

Biology and Life Sciences

Physical Sciences

Humanities and Social Sciences

Practical and Applied Sciences

Memes

Miscellaneous

founded 3 years ago
MODERATORS
you are viewing a single comment's thread
view the rest of the comments
[–] The_Decryptor@aussie.zone 7 points 3 hours ago (2 children)

Ok, but who is making those "open weight" models though? Individuals don't really have the resources to run these huge scraping operations, so they're often still corporate releases with fake open source branding.

[–] percent@infosec.pub 1 points 4 minutes ago

There are huge public datasets that are often used for pretraining. Common Crawl and C4 are probably the most prominent, but there are others.

There are also big public datasets available for fine-running and instruction tuning.

The open weight models are getting pretty powerful, thanks to some Chinese labs.

[–] Grimy@lemmy.world 2 points 2 hours ago* (last edited 2 hours ago)

They come from corporate but you can at least run them without any kind of analytics or censorship, as well as fine tune them on consumer hardware.

Consumers aren't in the best position right now though, especially with the price hikes.