Rethinking Critic Learning in PPO: Understanding and Mitigating Value Flattening Paper • 2609.18708 • Published 13 days ago • 82
Memory as Plans: World-Action Modeling with Memory-Grounded Planning Paper • 2609.11561 • Published 19 days ago • 41
Scaling Automatic Research Agents via World Models Paper • 2608.12564 • Published about 1 month ago • 482
Marigold V2: Revisiting Diffusion Transformers for Monocular Depth Estimation Paper • 2609.08084 • Published 21 days ago • 72
CoVeR: Coverage-Based Token Pruning for Multi-View 3D Reasoning in VLMs Paper • 2609.08345 • Published 21 days ago • 24
Percolation Dynamics in Optimization : Variance Cascades and Discrete Scale Invariance Paper • 2609.02373 • Published 27 days ago • 14
Knowing When Not to Reuse: Conditional Experience Transfer in Autonomous LLM Post-Training Paper • 2608.26730 • Published Aug 27 • 155
DavidAU/Qwen3.8-27B-Cold-Fusion-GAIN-V1.1-NM-DAU-NEO-MAX-MTP-GGUF Image-Text-to-Text • 27B • Updated 12 days ago • 517k • 324
OpenART: Scaling Agent Red Teaming via Open-Ended Environment Evolution Paper • 2608.00677 • Published Aug 1 • 265
PAST-Bench: Benchmarking the Foundations of Recursive Self-Improvement in Personal Agents Paper • 2608.04003 • Published Aug 4 • 36
SkillJack: Persistent Skill Backdoors in Self-Evolving Agents Paper • 2608.03509 • Published Aug 4 • 24