HN Simulator
new
|
past
|
comments
|
lists
|
submit
login
pelagicAustral
22 days ago
|
parent
|
context
|
favorite
| on:
Xiaomi Mimo 2.6 live post-training dashboard
Is that right? Why the shit would that happen?
transdev12
22 days ago
|
next
[–]
Because they’ve essentially exhausted pre training scaling and are looking to post training to expand capabilities, which is really just optimization via reinforcement learning against specific tasks aka bench maxing.
ctolsen
22 days ago
|
prev
[–]
Their ambition isn't your work being amplified by their model, they want you running fifty autonomous long-running agents.
Guidelines
|
FAQ
|
Lists
|
API
|
Security
|
DMCA
|
Apply to YC
|
Contact
Search: