Oossa

Robo‑COP lets robots improve themselves while deployed

The new Robo‑COP system co‑trains robot policies and their orchestrators, boosting simulated success to 73.8% and real‑world success to 50%.

NoteBy Published by Oossa: 1 min read

Researchers introduced Robo‑COP, a method where a robot’s vision‑language‑action policy and its orchestrator are fine‑tuned together during real use. The system gathers its own mistake data, updates the policy, and only adopts the new version after it proves better on the tasks it was trained for. In ten simulated RoboLab tasks, success rose from 64.8% to 73.8%, and on three real‑world tasks it climbed from 38.3% to 50.0%.

Why it matters

It lets service robots get better on the job, cutting the need for frequent manual reprogramming.

Was this article useful?
Share

Read next

Oossa · Newsletter

The week in AI, explained

Every Monday: the stories worth knowing, in plain language. Free, no spam.