Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

You probably meant Qwen3.8-35B-A3B. But judging from some of the words from their team, it seems unlikely unfortunately.


They normally release a 35b dense and an 27b moe (4B active per token)

For context 35B on my m4 runs at 10 tokens a second, 27B moe runs 50-60 tokens a second.


You have your numbers switched. 27B is the dense model and runs slowly on unified memory. 35B (A3B active) runs great on unified memory.


27B dense or 35B-A3B MoE. You might be confusing it with Gemma 4 that has a 26B-A4B variant.




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: