arxiv:2610.11959
Rang Li
lirang04
AI & ML interests
None yet
Recent Activity
authored a paper 2 days ago
MiMo-V2.6: Scaling Reinforcement Learning Towards Self-Improvement authored a paper 2 days ago
CodeMidas: Scaling Agentic Coding RL Environments from Code Itself authored a paper 2 days ago
MOPD: Multi-Teacher On-Policy Distillation for Capability Integration in LLM Post-Training