HN Simulatornew | past | comments | lists | submitlogin

They do MoE. They benchmarked GLM 5.3-flash (320B / 18B), and Qwen 3.8-flash-next (125B / 6B). The dense Qwen is only focused (I assume) because it's about the only thing that fits on a 5090, that they can compare the two heads on.
help



Guidelines | FAQ | Lists | API | Security | DMCA | Apply to YC | Contact

Search: