this post was submitted on 29 Sep 2026
170 points (96.2% liked)
Technology
88390 readers
3362 users here now
This is a most excellent place for technology news and articles.
Our Rules
- Follow the lemmy.world rules.
- Only tech related news or articles.
- Be excellent to each other!
- Mod approved content bots can post up to 10 articles per day.
- Threads asking for personal tech support may be deleted.
- Politics threads may be removed.
- No memes allowed as posts, OK to post as comments.
- Only approved bots from the list below, this includes using AI responses and summaries. To ask if your bot can be added please contact a mod.
- Check for duplicates before posting, duplicates may be removed
- Accounts 7 days and younger will have their posts automatically removed.
Approved Bots
founded 3 years ago
MODERATORS
you are viewing a single comment's thread
view the rest of the comments
view the rest of the comments
😂 Some of this sounds like a child wrote it (emphasis mine):
Well you would be given that no one else calls it that.
But also, have you heard of Qwen? Last time I checked it was as good if not better than American models! (I admit I could be wrong about that now as I haven't checked benchmarks in a few months now)
Well, it’s quite easy being the world leader in your just-made-up field
It's definitely better than American open-weight models. The gap between American flagship models and Chinese open-weight models is closing quickly though.
With the token prices of Chinese models, I seldom reach for American models anymore.
I've spent the last week or so coding an application using nothing but Qwen3.8-27B. I was curious how much I could do entirely with a local model running on my own personal hardware, and it turns out the answer was "everything."
Granted, it's not the fanciest application. But it's fun.
Yep, I've been using that same model pretty heavily too, running on ExLlamaV3. It's very impressive for such a little model!
I'm really looking forward to seeing what Qwen4 is like. 3.8 was just 3.6 with additional training, my understanding is that Qwen4 is new from the ground up.
I hope they make a MoE version. I get like 50 tokens/sec with Qwen3.6-35B-A3B, but can't rely on it quite as much as Qwen3.8-27B.
Sadly they don't seem to have 36B-A3B models on their roadmap any more, they didn't do one for 3.8 either. I agree it was a nice sweet spot between speed and capability, I still use the 3.6 version of 36B-A3B for larger-scale local work. Maybe someone else will aim for that. Or Qwen4 will do something new with the architecture that makes it unnecessary.
My Hermes instance periodically checks the overall open-weight LLM landscape, and it recently recommended trying Ornith 1.5 35B-A3B.
I haven't tried it yet, but on paper, it sounds like it has some potential.
Heh, I downloaded that one just recently, I read that it was good at natural prose and I've been working on a little pet project to make a framework for auto-writing short stories based on a simple premise. Haven't tested it extensively yet though.