Hacker News
new
|
past
|
comments
|
ask
|
show
|
jobs
|
submit
login
WiSaGaN
30 days ago
|
parent
|
context
|
favorite
| on:
Qwen 3.8-Flash-Next releasing tomorrow (125B a6B)
You probably meant Qwen3.8-35B-A3B. But judging from some of the words from their team, it seems unlikely unfortunately.
vorticalbox
30 days ago
[–]
They normally release a 35b dense and an 27b moe (4B active per token)
For context 35B on my m4 runs at 10 tokens a second, 27B moe runs 50-60 tokens a second.
cpburns2009
29 days ago
|
parent
|
next
[–]
You have your numbers switched. 27B is the dense model and runs slowly on unified memory. 35B (A3B active) runs great on unified memory.
yencabulator
29 days ago
|
parent
|
prev
[–]
27B dense or 35B-A3B MoE. You might be confusing it with Gemma 4 that has a 26B-A4B variant.
Guidelines
|
FAQ
|
Lists
|
API
|
Security
|
Legal
|
Apply to YC
|
Contact
Search: