Our models are trained on top of the UI-TARS-1.5-7B model using **ACuRL**, an **A**utonomous **Cu**rriculum **R**einforcement **L**earning framework that steers agents to continually learn in target environments with zero human data. For more details, please refer to [here](https://github.com/OSU-NLP-Group/ACuRL)