Papers › xLSTM-SENet: xLSTM for Single-Channel Speech Enhancement
xLSTM-SENet: xLSTM for Single-Channel Speech Enhancement
Nikolai Lund Kühne, Jan Østergaard, Jesper Jensen, Zheng-Hua Tan
While attention-based architectures, such as Conformers, excel in speech enhancement, they face challenges such as scalability with respect to input sequence length. In contrast, the recently proposed Extended Long Short-Term Memory (xLSTM) architecture offers linear scalability. However, xLSTM-based models remain unexplored for speech enhancement. This paper introduces xLSTM-SENet, the first xLSTM-based single-channel speech enhancement system. A comparative analysis reveals that xLSTM-and notably, even LSTM-can match or outperform state-of-the-art Mamba- and Conformer-based systems across various model sizes in speech enhancement on the VoiceBank+Demand dataset. Through ablation studies, we identify key architectural design choices such as exponential gating and bidirectionality contributing to its effectiveness. Our best xLSTM-based model, xLSTM-SENet2, outperforms state-of-the-art Mamba- and Conformer-based systems of similar complexity on the Voicebank+DEMAND dataset.
Code
Repository list and official/mentioned flags are the archive's, frozen 2025-07-28. Reachability, where shown, is from one Syntology probe window (2026-09-16 to 2026-09-18); repositories not probed show nothing. GitHub stars are not tracked.
Code Syntology ran Syntology
Not run by Syntology. Nothing on this page verifies that the listed code works.
Tasks
Results from the paper archive 2025-07-28
| Task | Dataset | Model | Metric | Value | Rank at snapshot | Leaderboard | Report |
|---|---|---|---|---|---|---|---|
| Speech Enhancement | VoiceBank + DEMAND | xLSTM-SENet2 | CBAK | 3.98 | #11 of 42 | Archive leaderboard | report |
| Speech Enhancement | VoiceBank + DEMAND | xLSTM-SENet2 | COVL | 4.27 | #11 of 42 | Archive leaderboard | report |
| Speech Enhancement | VoiceBank + DEMAND | xLSTM-SENet2 | CSIG | 4.78 | #11 of 42 | Archive leaderboard | report |
| Speech Enhancement | VoiceBank + DEMAND | xLSTM-SENet2 | PESQ (wb) | 3.53 | #11 of 42 | Archive leaderboard | report |
| Speech Enhancement | VoiceBank + DEMAND | xLSTM-SENet2 | Para. (M) | 2.27 | #11 of 42 | Archive leaderboard | report |
| Speech Enhancement | VoiceBank + DEMAND | xLSTM-SENet2 | STOI | 0.96 | #11 of 42 | Archive leaderboard | report |
Ranks are positions in the archive's leaderboards as they stood at the 2025-07-28 snapshot. Results published since then are not among these rows, so a rank here is not a current standing.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections