2026-09-23T20:01:36.188Z
- title:
G-Mamba: Sparse Graph-Guided Mamba for Audio-Visual Speech Enhancement→ SG-Mamba: Sparse Graph-Guided Mamba for Audio-Visual Speech Enhancement
SOURCE-LINKED INTELLIGENCE
Lightweight audio-visual speech enhancement (AVSE) models face a critical trade-off between computational efficiency and cross-modal alignment accuracy. While simple concatenation lacks relational expressiveness, dense cross-attention incurs computational overhead and is prone to unreliable cross-modal correspondence under strong acoustic interference. We propose Sparse Graph-Guided Mamba (SG-Mamba), a lightweight AVSE framework that integrates a sparse heterogeneous graph with a linear-complexity Mamba backbone. The graph explicitly models modality-specific relations through content-adaptive
Read original source ↗ Open in workspace
First collected: 2026-09-20T08:20:57.646Z. This is not the publication date.
AIIC observation times, not verified publisher revision times. Up to eight recent revisions.