Efficient Agent Training for Computer Use

22/05/2025 23 min Episodio 782

Listen "Efficient Agent Training for Computer Use"

Descargar episodio Ver en sitio original

Episode Synopsis

🤗 Upvotes: 32 | cs.AI, cs.CL, cs.LG

Authors:
Yanheng He, Jiahe Jin, Pengfei Liu

Title:
Efficient Agent Training for Computer Use

Arxiv:
http://arxiv.org/abs/2505.13909v1

Abstract:
Scaling up high-quality trajectory data has long been a critical bottleneck for developing human-like computer use agents. We introduce PC Agent-E, an efficient agent training framework that significantly reduces reliance on large-scale human demonstrations. Starting with just 312 human-annotated computer use trajectories, we further improved data quality by synthesizing diverse action decisions with Claude 3.7 Sonnet. Trained on these enriched trajectories, our PC Agent-E model achieved a remarkable 141% relative improvement, surpassing the strong Claude 3.7 Sonnet with extended thinking on WindowsAgentArena-V2, an improved benchmark we also released. Furthermore, PC Agent-E demonstrates strong generalizability to different operating systems on OSWorld. Our findings suggest that strong computer use capabilities can be stimulated from a small amount of high-quality trajectory data.

More episodes of the podcast Daily Paper Cast

Native Parallel Reasoner: Reasoning in Parallelism via Self-Distilled Reinforcement Learning 09/12/2025

Beyond Real: Imaginary Extension of Rotary Position Embeddings for Long-Context LLMs 09/12/2025

Unified Video Editing with Temporal Reasoner 09/12/2025

Voxify3D: Pixel Art Meets Volumetric Rendering 09/12/2025

Scaling Zero-Shot Reference-to-Video Generation 09/12/2025

DoVer: Intervention-Driven Auto Debugging for LLM Multi-Agent Systems 09/12/2025

TwinFlow: Realizing One-step Generation on Large Models with Self-adversarial Flows 08/12/2025

EditThinker: Unlocking Iterative Reasoning for Any Image Editor 08/12/2025

From Imitation to Discrimination: Toward A Generalized Curriculum Advantage Mechanism Enhancing Cross-Domain Reasoning Tasks 08/12/2025

EMMA: Efficient Multimodal Understanding, Generation, and Editing with a Unified Architecture 08/12/2025

Ver todos los episodios

ZARZA We are Zarza, the prestigious firm behind major projects in information technology.

Efficient Agent Training for Computer Use

Listen "Efficient Agent Training for Computer Use"

Episode Synopsis

More episodes of the podcast Daily Paper Cast

Googling with breathtaking tricks you ignore

Gray Hat Hacking, those with ambiguous ethics…

Bandwidth: Broadband or Narrowband?

Personnel recruitment via Web

Deep web or Invisible Internet

Subdomains, a glance with the experts!

Free Internet, a prediction in Nostradamus style

Educational Technology: From traditional to digital

Localhost, there’s no place like 127.0.0.1

Googling with breathtaking tricks you ignore

Gray Hat Hacking, those with ambiguous ethics…

Internet Predators on the prowl

Dot COM: The Internet’s dominant TLD