Generalized Agent Iteration: One Formal Framework for Iterative Policy Improvement and Recursive Self-Improvement Paper • 2609.13406 • Published 7 days ago • 59
Occamy-1.0: Open Pareto-frontier 35B Intelligence for Co-work Paper • 2609.11977 • Published 14 days ago • 126
ZGCM-1: A Fully Open and Extremely Efficient Foundation Model for Math and Agentic Search Paper • 2609.13356 • Published 7 days ago • 309
Dream-RSI: Recursive Self-Improvement through Evolving Worlds Paper • 2609.14858 • Published 4 days ago • 290
DataFlex-RL: An Evaluation Platform for RLVR Data Policies Paper • 2609.06107 • Published 13 days ago • 125
TANGO: Humanoid Navigation in Cluttered Environments with a Whole-Body Vision-Language-Action Model Paper • 2609.09158 • Published 10 days ago • 23
One Editor, Many Edits: A Unified Training-Free Framework for Diverse Video Editing Paper • 2609.04190 • Published 15 days ago • 8
NeoHorse-1: Towards Recursive Self-Improvement via Agentic Post-Training with Routing Harness Paper • 2609.08183 • Published 10 days ago • 420