WorldAuditBench: Interactive 3D World Auditing with Multimodal Agents Paper • 2609.40325 • Published 9 days ago • 104
Mid-Harness: Scaling Actions Between Model and Harness for Terminal Agents Paper • 2609.39982 • Published 9 days ago • 119
The Teacher Is a Direction, Not a Destination: Extrapolating RL-Induced Representation Residuals in On-Policy Distillation Paper • 2609.36484 • Published 10 days ago • 582
LEGO-Anything: Coding Agents for 3D Scene Reconstruction Paper • 2609.36380 • Published 11 days ago • 140
PanoVLN: Towards Effective Panoramic Vision-and-Language Navigation Paper • 2609.34759 • Published 11 days ago • 168
FlashRender: Few-Step Generative Rendering via Camera-Controlled Video MeanFlow Paper • 2609.03563 • Published Sep 3 • 18
WorldReward: Reward Modeling for Camera-Conditioned World Models Paper • 2609.03952 • Published Sep 3 • 28
Why Gated DeltaNet Survives 4-Bit Quantization: NVFP4 W4A4 for the Recurrent Half of a Hybrid 27B LLM Paper • 2609.04098 • Published Sep 3 • 86
Build error Agents Featured 563 Talking Face Generation with Multilingual TTS 👄 563 Generate multilingual talking-face videos from your text
Efficient Stable Diffusion Collection Block-removed Knowledge-distilled SD models; https://github.com/Nota-NetsPresso/BK-SDM • 9 items • Updated Jul 1, 2024 • 6
Efficient Solar Open Family Collection Solar Open models officially optimized by Nota AI for Korea’s government-led Sovereign AI initiative as part of the Upstage consortium. • 7 items • Updated Aug 31 • 26
Out of Sight but Not Out of Mind: Hybrid Memory for Dynamic Video World Models Paper • 2603.25716 • Published Mar 26 • 76