Learn What's Left, Not What's Mastered: Saturation Aware Advantage Reweighting for Multi-Reward Policy Optimization Paper • 2608.16072 • Published 3 days ago • 141
Running Featured 694 Agent Memory Leaderboard 🧠 694 Unified memory evaluation · Results expected August 12.
Qwen-UI-Agent Technical Report: Toward Next-Generation Real-World Centric Foundation GUI Agents Paper • 2607.28227 • Published 21 days ago • 306
ABot-World-0: Infinite Interactive World Rollout on a Single Desktop GPU Paper • 2607.19191 • Published 29 days ago • 311
AREX: Towards a Recursively Self-Improving Agent for Deep Research Paper • 2607.21461 • Published 28 days ago • 154
PalmClaw: A Native On-Device Agent Framework for Mobile Phones Paper • 2607.13027 • Published Jul 14 • 15
LongStraw: Long-Context RL Beyond 2M Tokens under a Fixed GPU Budget Paper • 2607.14952 • Published Jul 16 • 212
WildCity: A Real-World City-Scale Testbed for Rendering, Simulation, and Spatial Intelligence Paper • 2607.06838 • Published Jul 7 • 14
OmniOpt: Taxonomy, Geometry, and Benchmarking of Modern Optimizers Paper • 2607.04033 • Published Jul 4 • 76
SkillHone: A Harness for Continual Agent Skill Evolution Through Persistent Decision History Paper • 2606.08671 • Published Jun 23 • 47
MemSlides: A Hierarchical Memory Driven Agent Framework for Personalized Slide Generation with Multi-turn Local Revision Paper • 2606.17162 • Published Jun 15 • 178
Skill-MAS: Evolving Meta-Skill for Automatic Multi-Agent Systems Paper • 2606.18837 • Published Jun 17 • 58