GTR: Gated Token Recurrence for Efficient Dense Prediction Paper • 2609.26590 • Published 18 days ago • 16
Qwen-Planner-Agent: A Closed-Loop AI-for-AI Framework for Real-World Mobile Planner Agents Paper • 2609.29892 • Published 17 days ago • 34
LatentPort: Beyond KV Cache - Cross-Model Transfer of Recurrent Memory in Hybrid Language Models: A 4B-to-9B Hybrid-State Handoff Without Target Prefix Replay Paper • 2609.25053 • Published Sep 7 • 17
Jev-Mem: System-One-Controlled Agentic Memory for Efficient AI Agents Paper • 2609.23986 • Published 20 days ago • 29
DeepSeek-V4.1-Flash: Pushing the Limits of KV Cache Compression Paper • 2609.19969 • Published 24 days ago • 228
Gaze as Evidence for Common Grounding: A Cross-Corpus Analysis of MapTask and MUNDEX Paper • 2609.18011 • Published 25 days ago • 30
Continual Learning Mechanisms Compose for Long-Horizon Memorization Paper • 2609.06986 • Published Sep 7 • 376
Mind2Dialogue: Training Human-Aware Language Models by Simulating User Mental States Paper • 2609.15972 • Published 27 days ago • 9
Co-Evolving Harnesses and Models: On-Policy Correction Helps Weaker Models Catch Up Where Imitation Fails Paper • 2609.09134 • Published Sep 8 • 11
Harnessing CLIP and DINO: An Uncertainty-Aware Cascaded Fusion Network for Generalizable Deepfake Image Detection Paper • 2609.07670 • Published Sep 7 • 18
Safety for Whom? Boundary-Aware Self-Distillation for Controlled LLM Safety Refusal Paper • 2609.04482 • Published Sep 3 • 11