Mind the Gap: Bridging Thought Leap for Improved Chain-of-Thought Tuning
xhl
zjuxhl
AI & ML interests
None yet
Recent Activity
upvoted a paper 3 days ago
Reflect, Revise, Reuse: Training-Free Skill Evolution for GUI Agents upvoted a paper 3 days ago
RetireOPD: Self-Retiring On-Policy Distillation for Agentic Reinforcement Learning updated a Space 11 days ago
zjuxhl/EasySteer