U-Space: Uncovering When and Why Uncertainty Arises in Language Models Paper • 2610.09087 • Published 6 days ago • 39
MC-Sparse: Deconstructing and Closing the Dense-Sparse Attention Gap in Diffusion Transformers Paper • 2610.06801 • Published 7 days ago • 38
MiMo-V2.6: Scaling Reinforcement Learning Towards Self-Improvement Paper • 2610.11959 • Published 4 days ago • 74
Mechanics of Long-Context Hybrid Models Part 1.1: From Hybrid Attention to Hybrid Position Paper • 2610.10114 • Published 5 days ago • 31
nanoMuse: An Open-Source Personal Agent for Every Device You Own Paper • 2610.08699 • Published 6 days ago • 101
In-Distribution Forcing for Long Video Generation at Test Time Paper • 2610.03120 • Published 10 days ago • 49
How to Loop MoE: Flatten the Experts, Untie the Attention Paper • 2609.35751 • Published 14 days ago • 13
Towards Looped Models Done Right, Part II: Rethinking at Fixed Points Paper • 2610.06833 • Published 7 days ago • 33
Memadapter: Counterfactual Adaptation Against Memory-induced Sycophancy Paper • 2610.05162 • Published 8 days ago • 60
VeriHarness: Scaling Agentic Verification for Long-Horizon Tasks Paper • 2610.00972 • Published 11 days ago • 61
AREX-2: Advancing Self-Improving Agents through Long-Horizon Reflective Tasks Paper • 2609.38288 • Published 13 days ago • 137
In-Context Learning for Robots: Methods and Applications Paper • 2609.36012 • Published 14 days ago • 322
FuseReg: Regularizing Layer Fusion Mitigates the Reconstruction-Generation Gap in Representation Autoencoders Paper • 2609.31620 • Published 17 days ago • 145