Canonical Cortical Graph Neural Networks and its Application for Speech Enhancement in Audio-Visual Hearing Aids

Passos, Leandro A.; Papa, João Paulo; Hussain, Amir; Adeel, Ahsan

doi:10.1016/j.neucom.2022.11.081

Computer Science > Sound

arXiv:2206.02671 (cs)

[Submitted on 6 Jun 2022 (v1), last revised 31 Jan 2023 (this version, v3)]

Title:Canonical Cortical Graph Neural Networks and its Application for Speech Enhancement in Audio-Visual Hearing Aids

Authors:Leandro A. Passos, João Paulo Papa, Amir Hussain, Ahsan Adeel

View PDF

Abstract:Despite the recent success of machine learning algorithms, most models face drawbacks when considering more complex tasks requiring interaction between different sources, such as multimodal input data and logical time sequences. On the other hand, the biological brain is highly sharpened in this sense, empowered to automatically manage and integrate such streams of information. In this context, this work draws inspiration from recent discoveries in brain cortical circuits to propose a more biologically plausible self-supervised machine learning approach. This combines multimodal information using intra-layer modulations together with Canonical Correlation Analysis, and a memory mechanism to keep track of temporal data, the overall approach termed Canonical Cortical Graph Neural networks. This is shown to outperform recent state-of-the-art models in terms of clean audio reconstruction and energy efficiency for a benchmark audio-visual speech dataset. The enhanced performance is demonstrated through a reduced and smother neuron firing rate distribution. suggesting that the proposed model is amenable for speech enhancement in future audio-visual hearing aid devices.

Subjects:	Sound (cs.SD); Computer Vision and Pattern Recognition (cs.CV); Audio and Speech Processing (eess.AS)
Cite as:	arXiv:2206.02671 [cs.SD]
	(or arXiv:2206.02671v3 [cs.SD] for this version)
	https://doi.org/10.48550/arXiv.2206.02671
Related DOI:	https://doi.org/10.1016/j.neucom.2022.11.081

Submission history

From: Leandro Passos [view email]
[v1] Mon, 6 Jun 2022 15:20:07 UTC (1,293 KB)
[v2] Tue, 22 Nov 2022 11:21:19 UTC (2,002 KB)
[v3] Tue, 31 Jan 2023 14:14:49 UTC (2,002 KB)

Computer Science > Sound

Title:Canonical Cortical Graph Neural Networks and its Application for Speech Enhancement in Audio-Visual Hearing Aids

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Sound

Title:Canonical Cortical Graph Neural Networks and its Application for Speech Enhancement in Audio-Visual Hearing Aids

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators