← Về trang chính1 tin
1
Xiaomi mở mã nguồn hơn 7.000 môi trường MiMo RL và code huấn luyện trên Hugging Face
Releases4 ngày trướcXiaomi vừa công khai mã nguồn hơn 7.000 môi trường MiMo RL cùng code huấn luyện trên Hugging Face, cho phép cộng đồng nghiên cứu và phát triển các agent reinforcement learning tại quy mô lớn.
Đăng nhập để góp ý, chỉnh sửa
Nguồn chính
- SOURCE@artificialanlys“it retains the same attractive pricing at $0.435 per 1M input tokens (with a 99% cache-hit discount) and $0.87 per 1M output tokens”x.com
- SOURCE@victormustar“Pro scores 46.32 on the Artificial Analysis Intelligence Index, the highest of any open model to date”x.com
- SOURCE@eliebakouch“xiaomi MiMo oss RL environments + training code”x.com
- SUPPORT@nrehiew_“The pro model went from 58.41 to 72.57 on DeepSWE”x.com
- SUPPORT@_luofuli“we have released a Qwen model distilled from MiMo RL trajectories as a stronger starting point for RL, along with 7K diverse environments and a complete RL training framework”x.com
- SUPPORT@teortaxestex“The team estimates a 10× productivity gain, shortening the R&D cycle from one month to 2–3 days”x.com
- SUPPORT@teortaxestex“This RL run had cost peanuts, 5 days of a ≈10K GPU cluster, it didn't even target CritPt explicitly”x.com
- SUPPORT@kimmonismus“Better than Grok 4.7, much cheaper and holy moly is china back. That is the real surprise!”x.com