HERO Collection "Doing More with Less Tokens: Hierarchical Reinforcement Learning for Efficient Coding Agents". Official 4B, 9B, and 35B-A3B models. • 3 items • Updated 2 days ago • 1
Active-SWE: Benchmarking Coding Agents for Proactive Bug Fixing without Issue Reports Paper • 2608.04682 • Published Aug 5 • 4