-
The first full-parameter 4B agent model to rank on 8 long-horizon and complex agent benchmarks, including GAIA, HLE, and BrowserComp, in the on-device setting.
-
Capable of over 100 rounds of continuous environment interaction, supporting multi-source information cross-validation, dynamic search strategy adjustment, and real-time verification of up-to-date information, enabling sustained deep exploration until task completion.
-
Fully open-sourced end-to-end, including (1) AgentRL, a fully asynchronous reinforcement learning framework for agent training, (2) AgentDock, a unified management and scheduling platform for tool sandboxes, (3) AgentToLeaP, a one-click evaluation platform for agent tool-learning capabilities. These components collectively support community collaboration and custom extensibility.