TokenRouter: Efficient Serving System for Token-Level LLM Routing Paper • 2610.12242 • Published 3 days ago • 130
MiMo-V2.6: Scaling Reinforcement Learning Towards Self-Improvement Paper • 2610.11959 • Published 3 days ago • 72
STEPQuant: When and Where Errors Matter in Delta-Rule Recurrent State Quantization Paper • 2609.38169 • Published 12 days ago • 118
Transformers Stop Thinking Too Early, and a Tiny LoRA Fixes It Paper • 2609.36585 • Published 12 days ago • 70
Kev Collection Jev-like family of decision models built on top of Qwen3.5/3.8 you can train and run on your own • 6 items • Updated 9 days ago • 33
FuseReg: Regularizing Layer Fusion Mitigates the Reconstruction-Generation Gap in Representation Autoencoders Paper • 2609.31620 • Published 16 days ago • 145
Holo4 Collection State-of-the-art Visual Language Model (VLM) for computer/mobile use, tool calls, and code. • 11 items • Updated 13 days ago • 42
Qwen3.8-3.6-27B Blend Collection Junie Local's unofficial JetBrains 50/50 blend of Qwen3.6-27B and Qwen3.8-27B: BF16, GGUF, and MLX weights. • 5 items • Updated 13 days ago • 9
OmniEcho: Spatial Audio Understanding for Embodied Agents Paper • 2609.23407 • Published 21 days ago • 35
CodeMidas: Scaling Agentic Coding RL Environments from Code Itself Paper • 2609.22068 • Published 23 days ago • 138
AuK Technical Report: An Open-Source Foundational Model for Speech Generation and Editing Paper • 2609.08936 • Published Sep 8 • 161