One Block, Multiple Depths: Recurrent Vision Transformers with Depth-Programmed Experts Paper • 2610.12448 • Published 4 days ago • 7
Arm-wise Compositional Generalization in Dual-Arm Vision-Language-Action Models Paper • 2610.06184 • Published 7 days ago • 5
ISTA-DASLab/Qwen3.8-Flash-Next-GSQ-RCO-Coder-GGUF Image-Text-to-Text • 117B • Updated 12 days ago • 679k • 376
Explore Broadly, Reason Sharply: Push Small Models toward the Frontier via Sampling Paper • 2609.38104 • Published 13 days ago • 10
Can MiniMax-H3 Reason About the Physical World? An Evaluation of Omni-Modal Generative Model Paper • 2609.18323 • Published 26 days ago • 120
Agora: Git as Shared Memory for Collective AutoResearch Paper • 2609.18094 • Published 26 days ago • 58
Think Before You Link: Rarity, Reasoning, and Retrieval in Multilingual Entity Linking Paper • 2609.10745 • Published Sep 9 • 22