arxiv:2603.02604
Yikun Ban
Yikunb
AI & ML interests
Reinforcement Learning
Recent Activity
upvoted a paper 3 days ago
Scaling Automatic Research Agents via World Models upvoted a paper about 1 month ago
SPOT: Sparse Probing and Outcome Calibration for On-Policy Distillation upvoted a paper about 1 month ago
Deferred Exposure of Future Trajectories for Verifiable Reasoning in Autonomous Driving VLMsOrganizations
None yet