Mind the Gap: Bridging Thought Leap for Improved Chain-of-Thought Tuning
xhl
zjuxhl
AI & ML interests
None yet
Recent Activity
upvoted a paper about 9 hours ago
Agent-G^2: Gaussian Guidance for Agentic Reinforcement Learning upvoted a paper about 9 hours ago
TTPO: Test-Time Policy Optimization authored a paper 28 days ago
Perceive-to-Reason: Decoupling Perception and Reasoning for Fine-Grained Visual Reasoning