HarnessEval-W: Agentifying the Evaluation of Visual Worlds Paper • 2608.16859 • Published 16 days ago • 339
Macaron-V1: Towards Open Continual Learning with Self-Improvement and Mixture-of-LoRA Paper • 2608.09819 • Published 23 days ago • 342
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Paper • 2607.29613 • Published Jul 31 • 28
Agentic Context Management: Solving Agent Memory and Cost by Treating Them as Lifecycle and Architecture Problems Paper • 2607.21503 • Published Jul 23 • 28
Embodied-R1.5: Evolving Physical Intelligence via Embodied Foundation Models Paper • 2606.11324 • Published Jun 9 • 172
The Flip Side of RLHF: On-Policy Feedback for Reward Model Self-Supervised Improvement Paper • 2605.30888 • Published May 29 • 11
JLT: Clean-Latent Prediction in Latent Diffusion Transformers Paper • 2605.27102 • Published May 26 • 33
Perception or Prejudice: Can MLLMs Go Beyond First Impressions of Personality? Paper • 2605.22109 • Published May 21 • 171
X-OmniClaw Technical Report: A Unified Mobile Agent for Multimodal Understanding and Interaction Paper • 2605.05765 • Published May 7 • 23
UniPool: A Globally Shared Expert Pool for Mixture-of-Experts Paper • 2605.06665 • Published May 7 • 13
Leveraging Verifier-Based Reinforcement Learning in Image Editing Paper • 2604.27505 • Published Apr 30 • 59
Rethinking Generalization in Reasoning SFT: A Conditional Analysis on Optimization, Data, and Model Capability Paper • 2604.06628 • Published Apr 8 • 330
GrandCode: Achieving Grandmaster Level in Competitive Programming via Agentic Reinforcement Learning Paper • 2604.02721 • Published Apr 3 • 640
Adam's Law: Textual Frequency Law on Large Language Models Paper • 2604.02176 • Published Apr 2 • 511