Score: 2

SEMNAV: A Semantic Segmentation-Driven Approach to Visual Semantic Navigation

Published: June 2, 2025 | arXiv ID: 2506.01418v1

By: Rafael Flor-Rodríguez , Carlos Gutiérrez-Álvarez , Francisco Javier Acevedo-Rodríguez and more

Potential Business Impact:

Helps robots find things using smart pictures.

Business Areas:

Navigation Navigation and Mapping

Visual Semantic Navigation (VSN) is a fundamental problem in robotics, where an agent must navigate toward a target object in an unknown environment, mainly using visual information. Most state-of-the-art VSN models are trained in simulation environments, where rendered scenes of the real world are used, at best. These approaches typically rely on raw RGB data from the virtual scenes, which limits their ability to generalize to real-world environments due to domain adaptation issues. To tackle this problem, in this work, we propose SEMNAV, a novel approach that leverages semantic segmentation as the main visual input representation of the environment to enhance the agent's perception and decision-making capabilities. By explicitly incorporating high-level semantic information, our model learns robust navigation policies that improve generalization across unseen environments, both in simulated and real world settings. We also introduce a newly curated dataset, i.e. the SEMNAV dataset, designed for training semantic segmentation-aware navigation models like SEMNAV. Our approach is evaluated extensively in both simulated environments and with real-world robotic platforms. Experimental results demonstrate that SEMNAV outperforms existing state-of-the-art VSN models, achieving higher success rates in the Habitat 2.0 simulation environment, using the HM3D dataset. Furthermore, our real-world experiments highlight the effectiveness of semantic segmentation in mitigating the sim-to-real gap, making our model a promising solution for practical VSN-based robotic applications. We release SEMNAV dataset, code and trained models at https://github.com/gramuah/semnav

WarNav: An Autonomous Driving Benchmark for Segmentation of Navigable Zones in War Scenes

CV and Pattern Recognition

Helps robots drive safely in war zones.

19 Nov 2025 0

88%

SemNav: A Model-Based Planner for Zero-Shot Object Goal Navigation Using Vision-Foundation Models

Robotics

Helps robots find things without prior training.

4 Jun 2025 1

88%

CityNavAgent: Aerial Vision-and-Language Navigation with Hierarchical Semantic Planning and Global Memory

Robotics

Drones follow spoken directions to fly in cities.

8 May 2025 1

View PDF Login to Bookmark

Country of Origin

🇪🇸 Spain

Repos / Data Links

github.com

Page Count

9 pages

SEMNAV: A Semantic Segmentation-Driven Approach to Visual Semantic Navigation

Helps robots find things using smart pictures.

Technical Abstract

WarNav: An Autonomous Driving Benchmark for Segmentation of Navigable Zones in War Scenes

SemNav: A Model-Based Planner for Zero-Shot Object Goal Navigation Using Vision-Foundation Models

CityNavAgent: Aerial Vision-and-Language Navigation with Hierarchical Semantic Planning and Global Memory