"Doing More with Less Tokens: Hierarchical Reinforcement Learning for Efficient Coding Agents". Official 4B, 9B, and 35B-A3B models.
XLearning-SCU
XLearning-SCU
AI & ML interests
None yet
Recent Activity
updated a collection about 18 hours ago
HERO updated a collection about 18 hours ago
HERO updated a collection about 18 hours ago
HEROOrganizations
None yet