OSReward: Instituting Standardized Evaluation for Cross-Platform Computer-Use Reward Models Paper ⢠2607.28609 ⢠Published 28 days ago ⢠73
OpenMobile: Building Open Mobile Agents with Task and Trajectory Synthesis Paper ⢠2604.15093 ⢠Published Apr 16 ⢠30
Intern-S1-Pro: Scientific Multimodal Foundation Model at Trillion Scale Paper ⢠2603.25040 ⢠Published Mar 26 ⢠134
TIDE: Trajectory-based Diagnostic Evaluation of Test-Time Improvement in LLM Agents Paper ⢠2602.02196 ⢠Published Feb 2 ⢠35
OSReward: Instituting Standardized Evaluation for Cross-Platform Computer-Use Reward Models Paper ⢠2607.28609 ⢠Published 28 days ago ⢠73
Diffusion of Thoughts: Chain-of-Thought Reasoning in Diffusion Language Models Paper ⢠2402.07754 ⢠Published Feb 12, 2024
MoS: Unleashing Parameter Efficiency of Low-Rank Adaptation with Mixture of Shards Paper ⢠2410.00938 ⢠Published Oct 1, 2024
OSReward: Instituting Standardized Evaluation for Cross-Platform Computer-Use Reward Models Paper ⢠2607.28609 ⢠Published 28 days ago ⢠73