DataFlex-RL: An Evaluation Platform for RLVR Data Policies Paper • 2609.06107 • Published 25 days ago • 165
Steering Geometry: Validating Human Value Geometry in LLM Steering Space Paper • 2609.06289 • Published 25 days ago • 31
UniWorld-View: Large-Baseline View Synthesis via Video Diffusion Models Paper • 2608.04701 • Published Aug 5 • 9
Qwen-UI-Agent Technical Report: Toward Next-Generation Real-World Centric Foundation GUI Agents Paper • 2607.28227 • Published Jul 30 • 158
Show, Don't Tell: Evaluating Spatial Cognition in Generative Pixels Rather Than LLM Text Paper • 2607.21072 • Published Jul 23 • 36
Progress Reward Modeling for Robotic Learning: A Comprehensive Survey Paper • 2607.21655 • Published Jul 22 • 111
ABot-World-0: Infinite Interactive World Rollout on a Single Desktop GPU Paper • 2607.19191 • Published Jul 21 • 114
Function-Aware Fill-in-the-Middle as Mid-Training for Coding Agent Foundation Models Paper • 2607.12463 • Published Jul 14 • 48
PalmClaw: A Native On-Device Agent Framework for Mobile Phones Paper • 2607.13027 • Published Jul 14 • 9
The Mirage of Optimizing Training Policies: Monotonic Inference Policies as the Real Objective for LLM Reinforcement Learning Paper • 2606.29526 • Published Jun 28 • 52
Walking in the Implicit: Interactive World Exploration via Neural Scene Representation Paper • 2606.30045 • Published Jun 29 • 8
Skill-MAS: Evolving Meta-Skill for Automatic Multi-Agent Systems Paper • 2606.18837 • Published Jun 17 • 11
DragMesh-2: Physically Plausible Dexterous Hand-Object Interaction with Articulated Objects Paper • 2606.15133 • Published Jun 13 • 34
Embodied-R1.5: Evolving Physical Intelligence via Embodied Foundation Models Paper • 2606.11324 • Published Jun 9 • 60
WeaveBench: A Long-Horizon, Real-World Benchmark for Computer-Use Agents with Hybrid Interfaces Paper • 2606.09426 • Published Jun 8 • 49
From Correctness to Utility: Gain-Based Prefix Evaluation for LLM Reasoning Paper • 2606.07190 • Published Jun 5 • 14
Imaginative Perception Tokens Enhance Spatial Reasoning in Multimodal Language Models Paper • 2606.03988 • Published Jun 3 • 26