Policy Management

ARPO

Official Implementation of ARPO: End-to-End Policy Optimization for GUI Agents with Experience Replay

Open Source

About

ARPO (Agentic Replay Policy Optimization) is designed to optimize policy for GUI agents using reinforcement learning. It focuses on completing long-horizon desktop tasks by processing multi-modal inputs, including screenshots and actions. The framework supports distributed rollouts for scalable task execution across parallel environments. It is particularly suited for developers and researchers working on GUI automation and reinforcement learning applications.

Open Source Health

Not enough history
Stars
163
Forks
11
License
Apache-2.0
Last commit
1 years ago
Python

Related Categories