# NavGPT-3 Sets New Benchmarks in Robot Navigation

> The system achieves human‑level success on indoor navigation tests, reaching 81.5% success on R2R‑CE and matching humans on RxR‑CE.

Oossa · 2026-10-09 · https://oossa.com/en/navgpt-3-sets-new-benchmarks-in-robot-navigation

Researchers led by Gengze Zhou introduced NavGPT-3, a runtime that links a large language model with a low‑latency action policy for robots. The harness runs reasoning, acting and monitoring as separate threads, letting the robot interrupt and switch tasks quickly. Their 8‑billion‑parameter action policy, trained on 19.28 million examples, already hit 74.5% success on the R2R‑CE benchmark. With the full system, NavGPT-3 improved to 81.5% on that test and matched human performance on RxR‑CE, completing routes in about 1 minute 22 seconds per episode.

## The facts

- Published on 2026-10-09
- NavGPT-3 achieves 81.51% success on R2R‑CE and 90.43% success on RxR‑CE

## Why it matters

It shows that robots can combine high‑level language reasoning with fast low‑level control, bringing autonomous navigation closer to human speed and reliability.

## Sources & references

1. [NavGPT-3: Harnessing Context in a Hierarchical Navigation Runtime](https://arxiv.org/abs/2610.10787) – arXiv cs.RO, 2026-10-09

Last updated: 2026-10-09
