WorldDirector: Building Controllable World Simulators with Persistent Dynamic Memory Paper • 2607.02517 • Published about 1 month ago • 33
The Mirage of Optimizing Training Policies: Monotonic Inference Policies as the Real Objective for LLM Reinforcement Learning Paper • 2606.29526 • Published Jun 28 • 170
ResearchStudio-Idea: An Evidence-Grounded Research-Ideation Skill Suite from ML Conference Outcomes Paper • 2607.04439 • Published 27 days ago • 63
ResearchStudio-Reel: Automate the Last Mile of Research from Paper to Poster, Video, and Blog Paper • 2607.04438 • Published 27 days ago • 64
UI-MOPD: Multi-Platform On-Policy Distillation for Continual GUI Agent Learning Paper • 2607.04425 • Published 27 days ago • 73
OmniOpt: Taxonomy, Geometry, and Benchmarking of Modern Optimizers Paper • 2607.04033 • Published 28 days ago • 76
SkillOpt-Lite: Better and Faster Agent Self-evolution via One Line of Vibe Paper • 2607.03451 • Published 29 days ago • 34
AlayaWorld: Long-Horizon and Playable Video World Generation Paper • 2607.06291 • Published 25 days ago • 91
Hierarchical Sparse Attention Done Right: Toward Infinite Context Modeling Paper • 2607.02980 • Published 29 days ago • 83
Accurate, Interdisciplinary and Transparent Structure-property Understanding with Deep Native Structural Reasoning Paper • 2607.07708 • Published 24 days ago • 88
Nemotron-Labs-3-Puzzle-75B-A9B: Compressing Hybrid MoE LLMs Paper • 2607.04371 • Published 27 days ago • 7
nvidia/NVIDIA-Nemotron-Labs-3-Puzzle-75B-A9B-NVFP4 Text Generation • 45B • Updated 25 days ago • 68.2k • 127
Dual Latent Memory in Vision-Language-Action Models for Robotic Manipulation Paper • 2607.07608 • Published 24 days ago • 57
Scaling Mixture-of-Experts Video Pretraining for Embodied Intelligence Paper • 2607.07675 • Published 24 days ago • 64
Program-as-Weights: A Programming Paradigm for Fuzzy Functions Paper • 2607.02512 • Published about 1 month ago • 238