Step 3.5 Flash: Open Frontier-Level Intelligence with 11B Active Parameters Paper • 2602.10604 • Published 15 days ago • 185
Code2World: A GUI World Model via Renderable Code Generation Paper • 2602.09856 • Published 16 days ago • 194
Recurrent-Depth VLA: Implicit Test-Time Compute Scaling of Vision-Language-Action Models via Latent Iterative Reasoning Paper • 2602.07845 • Published 19 days ago • 69
MOVA: Towards Scalable and Synchronized Video-Audio Generation Paper • 2602.08794 • Published 17 days ago • 154
DreamDojo: A Generalist Robot World Model from Large-Scale Human Videos Paper • 2602.06949 • Published 20 days ago • 35
On the Entropy Dynamics in Reinforcement Fine-Tuning of Large Language Models Paper • 2602.03392 • Published 23 days ago • 53
RISE-Video: Can Video Generators Decode Implicit World Rules? Paper • 2602.05986 • Published 21 days ago • 26
WideSeek-R1: Exploring Width Scaling for Broad Information Seeking via Multi-Agent Reinforcement Learning Paper • 2602.04634 • Published 22 days ago • 93
3D-Aware Implicit Motion Control for View-Adaptive Human Video Generation Paper • 2602.03796 • Published 23 days ago • 58
Green-VLA: Staged Vision-Language-Action Model for Generalist Robots Paper • 2602.00919 • Published 26 days ago • 305
GDPO: Group reward-Decoupled Normalization Policy Optimization for Multi-reward RL Optimization Paper • 2601.05242 • Published Jan 8 • 228
PaperBanana: Automating Academic Illustration for AI Scientists Paper • 2601.23265 • Published 27 days ago • 200
TTCS: Test-Time Curriculum Synthesis for Self-Evolving Paper • 2601.22628 • Published 28 days ago • 35