AcousticFusion: Fusing Sound Source Localization to Visual SLAM in   Dynamic Environments

Tianwei Zhang; Huayan Zhang; Xiaofei Li; Junfeng Chen; Tin Lun Lam and; Sethu Vijayakumar

arXiv:2108.01246·cs.RO·August 4, 2021

AcousticFusion: Fusing Sound Source Localization to Visual SLAM in Dynamic Environments

Tianwei Zhang, Huayan Zhang, Xiaofei Li, Junfeng Chen, Tin Lun Lam and, Sethu Vijayakumar

PDF

Open Access

TL;DR

This paper introduces AcousticFusion, a low-resource audio-visual fusion method that enhances SLAM in dynamic environments by integrating sound source direction, effectively removing dynamic obstacles and improving localization stability.

Contribution

It presents a novel fusion approach that combines sound source localization with visual SLAM, reducing computational costs and handling dynamic obstacles more efficiently.

Findings

01

Stable self-localization in dynamic environments

02

Low computational resource usage

03

Effective removal of dynamic obstacles

Abstract

Dynamic objects in the environment, such as people and other agents, lead to challenges for existing simultaneous localization and mapping (SLAM) approaches. To deal with dynamic environments, computer vision researchers usually apply some learning-based object detectors to remove these dynamic objects. However, these object detectors are computationally too expensive for mobile robot on-board processing. In practical applications, these objects output noisy sounds that can be effectively detected by on-board sound source localization. The directional information of the sound source object can be efficiently obtained by direction of sound arrival (DoA) estimation, but depth estimation is difficult. Therefore, in this paper, we propose a novel audio-visual fusion approach that fuses sound source direction into the RGB-D image and thus removes the effect of dynamic obstacles on the…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsIndoor and Outdoor Localization Technologies · Robotics and Sensor-Based Localization · Speech and Audio Processing