nanoMuse: An Open-Source Personal Agent for Every Device You Own Paper • 2610.08699 • Published 6 days ago • 101
MiMo-V2.6: Scaling Reinforcement Learning Towards Self-Improvement Paper • 2610.11959 • Published 4 days ago • 73
OneSearch-VL: Unified Multimodal Deep Research Agent for Image and Video Paper • 2610.12419 • Published 4 days ago • 24
OneStreamer: Unifying Perception, Memory, and Proactive Response in Streaming Video Interaction Paper • 2610.01762 • Published 11 days ago • 236
In-Context Learning for Robots: Methods and Applications Paper • 2609.36012 • Published 14 days ago • 322
A New Role for Relevance: Guiding Corpus Interaction in Agentic Search Paper • 2607.24223 • Published Jul 27 • 97
TimeLens2: Generalist Video Temporal Grounding with Multimodal LLMs Paper • 2607.17423 • Published Jul 19 • 101
Nemotron-Pre-Training-Datasets Collection Large scale pre-training datasets used in the Nemotron family of models. • 15 items • Updated Aug 11 • 194
JoyAI-VL-Interaction: Real-Time Vision-Language Interaction Intelligence Paper • 2606.14777 • Published Jun 10 • 220
Echo-Infinity: Learning Evolving Memory for Real-Time Infinite Video Generation Paper • 2606.04527 • Published Jun 3 • 29
ParaVT: Taming the Tool Prior Paradox for Parallel Tool Use in Agentic Video Reinforcement Learning Paper • 2605.20342 • Published May 19 • 31
SenseNova-U1: Unifying Multimodal Understanding and Generation with NEO-unify Architecture Paper • 2605.12500 • Published May 12 • 199
MiA-Signature: Approximating Global Activation for Long-Context Understanding Paper • 2605.06416 • Published May 7 • 56
Claw-Eval-Live: A Live Agent Benchmark for Evolving Real-World Workflows Paper • 2604.28139 • Published Apr 30 • 42