# New robot learning method ACG-WAM beats prior models on RoboTwin tasks

> The ACG-WAM approach reaches over 93% success in simulated scenes and 85% on real robots, surpassing the previous best by up to 10 points.

Oossa · 2026-10-07 · https://oossa.com/en/new-robot-learning-method-acg-wam-beats-prior-models-on-robotwin-tasks

Researchers have released a new method called ACG-WAM for teaching robots how to act based on visual input. In a set of 50 simulated RoboTwin 2.0 tasks, the system succeeded 93.46% of the time in clean environments and 92.68% when the scenes were randomized. On three real‑world tasks, it achieved 85.00% full success and a 91.67% partial‑completion score, beating the prior best system, Motus, by 10 and 9.17 percentage points respectively.

## How ACG-WAM works

The method adds an auxiliary predictor that learns to forecast geometric features—like object positions and shapes—several steps ahead, using the robot’s current camera view and the actions it will take. It trains this predictor with targets taken from a frozen visual encoder that processes pairs of current and future images. After training, the extra predictor and teacher modules are removed, leaving a lean model that can run on the robot.

## What this means for robot users

Higher success rates in both simulation and real hardware suggest that robots using ACG-WAM could learn tasks more reliably with less trial‑and‑error. For companies that deploy robots in factories or warehouses, the method could reduce downtime caused by failed attempts and shorten the time needed to teach new behaviors.

## The facts

- The paper was submitted on 3 Oct 2026 and posted on arXiv on 7 Oct 2026.
- ACG-WAM achieved 93.46% success in clean simulated RoboTwin 2.0 tasks.
- In randomized simulated scenes, it reached 92.68% success, the highest among compared methods.
- On three real‑robot tasks, it scored 85.00% full success and 91.67% partial completion.
- It outperformed the previous best system, Motus, by 10.00 and 9.17 percentage points respectively.

## Why it matters

For a factory manager, the higher success numbers mean robots can be taught new jobs faster and with fewer mistakes, potentially cutting training costs. However, the paper does not report long‑term reliability or how the method scales to more complex tasks, so those benefits remain to be confirmed in larger deployments.

## Sources & references

1. [ACG-WAM: World-Action Modeling via Action-Conditioned Geometric Latent Prediction](https://arxiv.org/abs/2610.06965) – arXiv cs.RO, 2026-10-07

Last updated: 2026-10-07
