Hacker News
new
|
past
|
comments
|
ask
|
show
|
jobs
|
submit
login
transdev12
7 days ago
|
parent
|
context
|
favorite
| on:
Xiaomi Mimo 2.6 live post-training dashboard
Because they’ve essentially exhausted pre training scaling and are looking to post training to expand capabilities, which is really just optimization via reinforcement learning against specific tasks aka bench maxing.
help
Guidelines
|
FAQ
|
Lists
|
API
|
Security
|
Legal
|
Apply to YC
|
Contact
Search: