-
Cut Your Losses! Learning to Prune Paths Early for Efficient Parallel Reasoning
Paper • 2604.16029 • Published • 23 -
Qwen3.5-Omni Technical Report
Paper • 2604.15804 • Published • 60 -
REFRAG: Rethinking RAG based Decoding
Paper • 2509.01092 • Published • 9 -
OneVL: One-Step Latent Reasoning and Planning with Vision-Language Explanation
Paper • 2604.18486 • Published • 96
Collections
Discover the best community collections!
Collections including paper arxiv:2604.16029
-
The Dragon Hatchling: The Missing Link between the Transformer and Models of the Brain
Paper • 2509.26507 • Published • 553 -
mHC: Manifold-Constrained Hyper-Connections
Paper • 2512.24880 • Published • 333 -
NeoVerse: Enhancing 4D World Model with in-the-wild Monocular Videos
Paper • 2601.00393 • Published • 133 -
LTX-2: Efficient Joint Audio-Visual Foundation Model
Paper • 2601.03233 • Published • 189
-
Contrastive Decoding Improves Reasoning in Large Language Models
Paper • 2309.09117 • Published • 40 -
Prometheus: Inducing Fine-grained Evaluation Capability in Language Models
Paper • 2310.08491 • Published • 57 -
Language Models are Hidden Reasoners: Unlocking Latent Reasoning Capabilities via Self-Rewarding
Paper • 2411.04282 • Published • 37 -
Insight-V: Exploring Long-Chain Visual Reasoning with Multimodal Large Language Models
Paper • 2411.14432 • Published • 26
-
PersonaVLM: Long-Term Personalized Multimodal LLMs
Paper • 2604.13074 • Published • 46 -
Elucidating the SNR-t Bias of Diffusion Probabilistic Models
Paper • 2604.16044 • Published • 72 -
Web Retrieval-Aware Chunking (W-RAC) for Efficient and Cost-Effective Retrieval-Augmented Generation Systems
Paper • 2604.04936 • Published • 26 -
Qwen3.5-Omni Technical Report
Paper • 2604.15804 • Published • 60
-
lusxvr/nanoVLM-222M
Image-Text-to-Text • 0.2B • Updated • 2.03k • 103 -
Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning
Paper • 2503.09516 • Published • 41 -
AlphaOne: Reasoning Models Thinking Slow and Fast at Test Time
Paper • 2505.24863 • Published • 98 -
QwenLong-L1: Towards Long-Context Large Reasoning Models with Reinforcement Learning
Paper • 2505.17667 • Published • 89
-
Cut Your Losses! Learning to Prune Paths Early for Efficient Parallel Reasoning
Paper • 2604.16029 • Published • 23 -
Qwen3.5-Omni Technical Report
Paper • 2604.15804 • Published • 60 -
REFRAG: Rethinking RAG based Decoding
Paper • 2509.01092 • Published • 9 -
OneVL: One-Step Latent Reasoning and Planning with Vision-Language Explanation
Paper • 2604.18486 • Published • 96
-
PersonaVLM: Long-Term Personalized Multimodal LLMs
Paper • 2604.13074 • Published • 46 -
Elucidating the SNR-t Bias of Diffusion Probabilistic Models
Paper • 2604.16044 • Published • 72 -
Web Retrieval-Aware Chunking (W-RAC) for Efficient and Cost-Effective Retrieval-Augmented Generation Systems
Paper • 2604.04936 • Published • 26 -
Qwen3.5-Omni Technical Report
Paper • 2604.15804 • Published • 60
-
The Dragon Hatchling: The Missing Link between the Transformer and Models of the Brain
Paper • 2509.26507 • Published • 553 -
mHC: Manifold-Constrained Hyper-Connections
Paper • 2512.24880 • Published • 333 -
NeoVerse: Enhancing 4D World Model with in-the-wild Monocular Videos
Paper • 2601.00393 • Published • 133 -
LTX-2: Efficient Joint Audio-Visual Foundation Model
Paper • 2601.03233 • Published • 189
-
lusxvr/nanoVLM-222M
Image-Text-to-Text • 0.2B • Updated • 2.03k • 103 -
Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning
Paper • 2503.09516 • Published • 41 -
AlphaOne: Reasoning Models Thinking Slow and Fast at Test Time
Paper • 2505.24863 • Published • 98 -
QwenLong-L1: Towards Long-Context Large Reasoning Models with Reinforcement Learning
Paper • 2505.17667 • Published • 89
-
Contrastive Decoding Improves Reasoning in Large Language Models
Paper • 2309.09117 • Published • 40 -
Prometheus: Inducing Fine-grained Evaluation Capability in Language Models
Paper • 2310.08491 • Published • 57 -
Language Models are Hidden Reasoners: Unlocking Latent Reasoning Capabilities via Self-Rewarding
Paper • 2411.04282 • Published • 37 -
Insight-V: Exploring Long-Chain Visual Reasoning with Multimodal Large Language Models
Paper • 2411.14432 • Published • 26