Switch language한국어
Back to the list

CUA-Gym: Scaling Verifiable Training Environments and Tasks for Computer-Use Agents

TL;DR AI

Key summary

2 min read
  1. Researchers introduced CUA-Gym, a scalable synthetic pipeline for training computer-use agents with verifiable rewards.

  2. The project produced 32,112 verified training tuples across 110 environments, including task instructions, environment states, and reward functions.

  3. Models trained on CUA-Gym improved performance on OSWorld-Verified and also transferred better to WebArena.

  4. The team plans to release the pipeline, environments, dataset, and models through CUA-Gym-Hub.

Read the original