Collections
Discover the best community collections!
Collections including paper arxiv:2608.07645
-
Multi-Agent Computer Use
Paper • 2606.01533 • Published • 8 -
OpenSkill: Open-World Self-Evolution for LLM Agents
Paper • 2606.06741 • Published • 30 -
Socratic-SWE: Self-Evolving Coding Agents via Trace-Derived Agent Skills
Paper • 2606.07412 • Published • 13 -
Bayesian-Agent: Posterior-Guided Skill Evolution for LLM Agent Harnesses
Paper • 2606.08348 • Published • 17
-
ReportBench: Evaluating Deep Research Agents via Academic Survey Tasks
Paper • 2508.15804 • Published • 15 -
StockBench: Can LLM Agents Trade Stocks Profitably In Real-world Markets?
Paper • 2510.02209 • Published • 57 -
Agent0: Unleashing Self-Evolving Agents from Zero Data via Tool-Integrated Reasoning
Paper • 2511.16043 • Published • 110 -
Agentic Reasoning for Large Language Models
Paper • 2601.12538 • Published • 208
-
EurekAgent: Agent Environment Engineering is All You Need For Autonomous Scientific Discovery
Paper • 2606.13662 • Published • 33 -
Toward Generalist Autonomous Research via Hypothesis-Tree Refinement
Paper • 2606.11926 • Published • 131 -
Retrospective Harness Optimization: Improving LLM Agents via Self-Preference over Trajectory Rollouts
Paper • 2606.05922 • Published • 72 -
AutoTrainess: Teaching Language Models to Improve Language Models Autonomously
Paper • 2606.31551 • Published • 24
-
EvoMaster: A Foundational Agent Framework for Building Evolving Autonomous Scientific Agents at Scale
Paper • 2604.17406 • Published • 7 -
Solvita: Enhancing Large Language Models for Competitive Programming via Agentic Evolution
Paper • 2605.15301 • Published • 23 -
Learning to Foresee: Unveiling the Unlocking Efficiency of On-Policy Distillation
Paper • 2605.11739 • Published • 61 -
SkillsVote: Lifecycle Governance of Agent Skills from Collection, Recommendation to Evolution
Paper • 2605.18401 • Published • 132
-
EurekAgent: Agent Environment Engineering is All You Need For Autonomous Scientific Discovery
Paper • 2606.13662 • Published • 33 -
Toward Generalist Autonomous Research via Hypothesis-Tree Refinement
Paper • 2606.11926 • Published • 131 -
Retrospective Harness Optimization: Improving LLM Agents via Self-Preference over Trajectory Rollouts
Paper • 2606.05922 • Published • 72 -
AutoTrainess: Teaching Language Models to Improve Language Models Autonomously
Paper • 2606.31551 • Published • 24
-
Multi-Agent Computer Use
Paper • 2606.01533 • Published • 8 -
OpenSkill: Open-World Self-Evolution for LLM Agents
Paper • 2606.06741 • Published • 30 -
Socratic-SWE: Self-Evolving Coding Agents via Trace-Derived Agent Skills
Paper • 2606.07412 • Published • 13 -
Bayesian-Agent: Posterior-Guided Skill Evolution for LLM Agent Harnesses
Paper • 2606.08348 • Published • 17
-
EvoMaster: A Foundational Agent Framework for Building Evolving Autonomous Scientific Agents at Scale
Paper • 2604.17406 • Published • 7 -
Solvita: Enhancing Large Language Models for Competitive Programming via Agentic Evolution
Paper • 2605.15301 • Published • 23 -
Learning to Foresee: Unveiling the Unlocking Efficiency of On-Policy Distillation
Paper • 2605.11739 • Published • 61 -
SkillsVote: Lifecycle Governance of Agent Skills from Collection, Recommendation to Evolution
Paper • 2605.18401 • Published • 132
-
ReportBench: Evaluating Deep Research Agents via Academic Survey Tasks
Paper • 2508.15804 • Published • 15 -
StockBench: Can LLM Agents Trade Stocks Profitably In Real-world Markets?
Paper • 2510.02209 • Published • 57 -
Agent0: Unleashing Self-Evolving Agents from Zero Data via Tool-Integrated Reasoning
Paper • 2511.16043 • Published • 110 -
Agentic Reasoning for Large Language Models
Paper • 2601.12538 • Published • 208