VABench: Measuring Embodied Spatial Intelligence through Visual Demonstrations, Active Perception, and Metric Control Paper • 2609.19554 • Published 15 days ago • 43
ProgramDistill: From Interactive Web Apps to Verifiable Reference-Guided SWE Tasks Paper • 2609.18805 • Published 16 days ago • 67
Continual Learning Mechanisms Compose for Long-Horizon Memorization Paper • 2609.06986 • Published 25 days ago • 376
RoboSPA: Can VLA Models Go Beyond Simple Scenes and Short-Horizon Tasks? Paper • 2609.05324 • Published 28 days ago • 28
deepseek-ai/DeepSeek-V4-Flash-Vision-Exp Image-Text-to-Text • 305B • Updated about 1 month ago • 966k • • 931
OpenART: Scaling Agent Red Teaming via Open-Ended Environment Evolution Paper • 2608.00677 • Published Aug 1 • 266
JigShape: Evaluating Visual-Geometric Reasoning in VLMs through Jigsaw Puzzles Paper • 2607.27670 • Published Aug 4 • 9
ContextMaster: Interactive Multi-Shot Video Creation via Fixed-Budget Sparse Context Routing Paper • 2608.04956 • Published Aug 5 • 17