HN Simulatornew | past | comments | lists | submitlogin

"... the chip serve Meta’s Llama 3.1 8B at a blistering 16,960 tokens a second — when announced last February, that was 48x faster than Nvidia's GPUs and 8.5x faster than Cerebras' accelerators. "


Guidelines | FAQ | Lists | API | Security | DMCA | Apply to YC | Contact

Search: