this post was submitted on 29 Sep 2026
170 points (96.2% liked)

Technology

88390 readers
3362 users here now

This is a most excellent place for technology news and articles.


Our Rules


  1. Follow the lemmy.world rules.
  2. Only tech related news or articles.
  3. Be excellent to each other!
  4. Mod approved content bots can post up to 10 articles per day.
  5. Threads asking for personal tech support may be deleted.
  6. Politics threads may be removed.
  7. No memes allowed as posts, OK to post as comments.
  8. Only approved bots from the list below, this includes using AI responses and summaries. To ask if your bot can be added please contact a mod.
  9. Check for duplicates before posting, duplicates may be removed
  10. Accounts 7 days and younger will have their posts automatically removed.

Approved Bots


founded 3 years ago
MODERATORS
you are viewing a single comment's thread
view the rest of the comments
[–] cman6@lemmy.world 19 points 2 days ago (2 children)

😂 Some of this sounds like a child wrote it (emphasis mine):

USHERING IN SUPER INTELLIGENCE: ... the executive branch to recognize the continuously advancing technological frontier and the limitless promise it offers the American people

RECOGNIZING THE IMMENSE OPPORTUNITY OF SI: The United States is the world leader in Super Intelligence...

Well you would be given that no one else calls it that.

But also, have you heard of Qwen? Last time I checked it was as good if not better than American models! (I admit I could be wrong about that now as I haven't checked benchmarks in a few months now)

[–] Photonic@lemmy.world 12 points 2 days ago

Well, it’s quite easy being the world leader in your just-made-up field

[–] percent@infosec.pub 4 points 2 days ago (1 children)

It's definitely better than American open-weight models. The gap between American flagship models and Chinese open-weight models is closing quickly though.

With the token prices of Chinese models, I seldom reach for American models anymore.

[–] FaceDeer@fedia.io 6 points 2 days ago (1 children)

I've spent the last week or so coding an application using nothing but Qwen3.8-27B. I was curious how much I could do entirely with a local model running on my own personal hardware, and it turns out the answer was "everything."

Granted, it's not the fanciest application. But it's fun.

[–] percent@infosec.pub 1 points 1 day ago (1 children)

Yep, I've been using that same model pretty heavily too, running on ExLlamaV3. It's very impressive for such a little model!

[–] FaceDeer@fedia.io 1 points 1 day ago (1 children)

I'm really looking forward to seeing what Qwen4 is like. 3.8 was just 3.6 with additional training, my understanding is that Qwen4 is new from the ground up.

[–] percent@infosec.pub 1 points 1 day ago* (last edited 1 day ago) (1 children)

I hope they make a MoE version. I get like 50 tokens/sec with Qwen3.6-35B-A3B, but can't rely on it quite as much as Qwen3.8-27B.

[–] FaceDeer@fedia.io 1 points 1 day ago (1 children)

Sadly they don't seem to have 36B-A3B models on their roadmap any more, they didn't do one for 3.8 either. I agree it was a nice sweet spot between speed and capability, I still use the 3.6 version of 36B-A3B for larger-scale local work. Maybe someone else will aim for that. Or Qwen4 will do something new with the architecture that makes it unnecessary.

[–] percent@infosec.pub 1 points 1 day ago (1 children)

My Hermes instance periodically checks the overall open-weight LLM landscape, and it recently recommended trying Ornith 1.5 35B-A3B.

I haven't tried it yet, but on paper, it sounds like it has some potential.

[–] FaceDeer@fedia.io 1 points 1 day ago

Heh, I downloaded that one just recently, I read that it was good at natural prose and I've been working on a little pet project to make a framework for auto-writing short stories based on a simple premise. Haven't tested it extensively yet though.