Canonical Cortical Graph Neural Networks and its Application for Speech   Enhancement in Audio-Visual Hearing Aids

Leandro A. Passos; Jo\~ao Paulo Papa; Amir Hussain; Ahsan Adeel

arXiv:2206.02671·cs.SD·February 1, 2023

Canonical Cortical Graph Neural Networks and its Application for Speech Enhancement in Audio-Visual Hearing Aids

Leandro A. Passos, Jo\~ao Paulo Papa, Amir Hussain, Ahsan Adeel

PDF

TL;DR

This paper introduces Canonical Cortical Graph Neural Networks, a biologically inspired model that effectively integrates multimodal data and temporal information, improving speech enhancement in audio-visual hearing aids with better accuracy and energy efficiency.

Contribution

It presents a novel biologically inspired GNN model combining intra-layer modulations, CCA, and memory mechanisms for multimodal and temporal data integration.

Findings

01

Outperforms state-of-the-art models in audio reconstruction

02

Reduces neuron firing rate for energy efficiency

03

Demonstrates potential for future hearing aid devices

Abstract

Despite the recent success of machine learning algorithms, most models face drawbacks when considering more complex tasks requiring interaction between different sources, such as multimodal input data and logical time sequences. On the other hand, the biological brain is highly sharpened in this sense, empowered to automatically manage and integrate such streams of information. In this context, this work draws inspiration from recent discoveries in brain cortical circuits to propose a more biologically plausible self-supervised machine learning approach. This combines multimodal information using intra-layer modulations together with Canonical Correlation Analysis, and a memory mechanism to keep track of temporal data, the overall approach termed Canonical Cortical Graph Neural networks. This is shown to outperform recent state-of-the-art models in terms of clean audio reconstruction…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.