Register and share your invite link to earn from video plays and referrals.

Bowen Wang
@BowenWangNLP
2nd year Ph.D. student @HKUniversity, Prev. @Tsinghua_Uni. Cooking digital agents at @Alibaba_Qwen, Prev @Kimi_moonshot
481 Following    1.1K Followers
RLVR has become the recipe for agentic post-training. But for Computer-Use Agents, the bottleneck is not the algorithm, it is the data. 🐌 🚀 We introduce CUA-Gym: a scalable, lightweight synthesis engine that turns arbitrary task queries into verifiable RLVR data for computer-use agents. The largest open CUA RLVR dataset to date: 🎯 32,122 verifiable RLVR tasks with programmatic setup scripts + rewards 🌐 110 environments: 16 desktop apps + 94 synthesized mock web apps 🏆 Qwen3.5-based CUA models trained with GSPO reach 72.6% on OSWorld-Verified and 56.6% on WebArena 📄 Paper: 🏠 Homepage: 🤗 Dataset: 💻 Codebase: 🧩 Environments: 🧵[1/6]
Show more