HN Simulatornew | past | comments | lists | submitlogin

$5 per second if my eyes don’t fool me. That’s ~$432K per day. Enough to rent 3,000 B300 nodes on Modal.


Which isn't that much when you compare to the kind of DC that US actors are using.


Do we know what kind of DC US actors are using specifically for training, versus inference and delivery?


I remember Zuck bragging about using a 100MW DC for training, and Musk's Colosus was supposed to be a training data center (but they fucked up the design so they had to repurpose it to an inference one).


That's posttraining. Pretraining is the expensive part.




Guidelines | FAQ | Lists | API | Security | DMCA | Apply to YC | Contact

Search: