TempCloze: Can Video-LLMs Identify the Missing Middle? Paper • 2609.01515 • Published 15 days ago • 32
ActReview: Rebuttal-Guided Training Data and Rubric Rewards for Actionable Peer Review Generation Paper • 2609.09076 • Published 8 days ago • 24
microsoft/VibeVoice-ASR-Streaming-7B Automatic Speech Recognition • 9B • Updated 12 days ago • 2.93k • 239
Knowing When Not to Reuse: Conditional Experience Transfer in Autonomous LLM Post-Training Paper • 2608.26730 • Published 20 days ago • 156
LLaDA-Image: Building Strong Image Generators with Fully Open Training Recipes Paper • 2609.03796 • Published 13 days ago • 236
GOAG: Generative and Object-Agnostic Grasp Planner for Dexterous Robotic Manipulation Paper • 2608.19759 • Published 27 days ago • 8
Can We Defend Against AI-Generated Video Attacks on Real-World Crisis Events? A Systematic Evaluation of Detectors, Generators and Social Dissemination Paper • 2608.14391 • Published Aug 14 • 282
CW-BASS v2: Saturation-Aware Pseudo-Label Selection for Semi-Supervised Segmentation under Foundation-Model Teachers Paper • 2608.12773 • Published Aug 13 • 9
Specification-first convergence with an AI coding agent: a case study of dismantling a core architectural invariant across 189 files in a 717k-line codebase with no test oracle and no human code review Paper • 2608.12440 • Published Aug 12 • 10
OpenART: Scaling Agent Red Teaming via Open-Ended Environment Evolution Paper • 2608.00677 • Published Aug 1 • 263
Relax Within, Balance Across: Geometry-Guided Load Balancing for Vision-Language Mixture-of-Experts Paper • 2608.00574 • Published Aug 1 • 9