Switch language한국어
Back to the list

Robostral Navigate

TL;DR AI

Key summary

2 min read
  1. Researchers introduced Robostral Navigate, an 8B vision-language navigation model that uses only monocular RGB images to predict waypoints in image space.

  2. Trained on 2.4 million simulated trajectories with prefix-caching and reinforcement learning, it improves training efficiency and cuts cost and time.

  3. The model sets state-of-the-art results on R2R-CE and RxR-CE, showing strong indoor navigation performance without expensive sensors or maps.

  4. Its RGB-only design makes navigation policies easier to deploy across different robot platforms and reduces recalibration overhead.

Read the original