Browse State-of-the-Art › Mamba
Mamba
552 papers with code · 0 benchmarks · 0 datasets archive 2025-07-28
Benchmarks archive 2025-07-28
No benchmark for this task in the archive.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
No dataset record in the archive lists this task.
Subtasks archive 2025-07-28
No subtask under this task in the archive's task tree.
Most implemented papers archive 2025-07-28
30 shown of 552 papers with code (1,058 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
1 Dec 2023 35 repositories listed Syntology ran 18 of 62 samples · 44 unverified · 28 pointer-only (licence)Foundation models, now powering most of the exciting applications in deep learning, are almost universally based on the Transformer architecture and its core attention module.
-
17 Jan 2024 15 repositories listed Syntology ran 2 of 13 samples · 11 unverified · 1 pointer-only (licence)The results demonstrate that Vim is capable of overcoming the computation & memory constraints on performing Transformer-style understanding for high-resolution images and it has great potential to be the…
-
18 Jan 2024 13 repositories listed Syntology ran 20 of 46 samples · 26 unverifiedAt the core of VMamba is a stack of Visual State-Space (VSS) blocks with the 2D Selective Scan (SS2D) module.
-
11 Dec 2023 6 repositories listed Syntology ran 2 of 2 samples · 0 unverifiedWhen used as a replacement for the standard attention layer in Transformers, the resulting gated linear attention (GLA) Transformer is found to perform competitively against the LLaMA-architecture Transformer (Touvron…
-
31 May 2024 5 repositories listed Syntology ran 6 of 13 samples · 7 unverified · 1 pointer-only (licence)While Transformers have been the main architecture behind deep learning's success in language modeling, state-space models (SSMs) such as Mamba have recently been shown to match or outperform Transformers at small to…
-
13 Jul 2024 4 repositories listedIt is too early to conclude that Mamba is a better alternative to transformers for speech before comparing Mamba with transformers in terms of both performance and efficiency in multiple speech-related tasks.
-
13 May 2024 4 repositories listedFor vision tasks, as image classification does not align with either characteristic, we hypothesize that Mamba is not necessary for this task; Detection and segmentation tasks are also not autoregressive, yet they…
-
5 Mar 2024 4 repositories listed Syntology ran 6 of 16 samples · 10 unverifiedLarge-scale sequence modeling has sparked rapid advances that now extend into biology and genomics.
-
29 Feb 2024 4 repositories listed Syntology ran 0 of 10 samples · 10 unverifiedRecurrent neural networks (RNNs) have fast inference and scale efficiently on long sequences, but they are difficult to train and hard to scale.
-
4 Feb 2024 4 repositories listed Syntology ran 11 of 17 samples · 6 unverified · 2 pointer-only (licence)To our best knowledge, this is the first medical image segmentation model constructed based on the pure SSM-based model.
-
17 Mar 2025 3 repositories listed Syntology ran 1 of 5 samples · 4 unverifiedRecent breakthroughs in solving reasoning, math and coding problems with Large Language Models (LLMs) have been enabled by investing substantial computation budgets at inference time.
-
2 Oct 2024 3 repositories listed Syntology ran 3 of 3 samples · 0 unverified · 2 pointer-only (licence)The introduction of Transformers in 2017 reshaped the landscape of deep learning.
-
10 Jul 2024 3 repositories listed Syntology ran 2 of 8 samples · 6 unverified · 6 pointer-only (licence)We propose a novel hybrid Mamba-Transformer backbone, MambaVision, specifically tailored for vision applications.
-
5 Jul 2024 3 repositories listed Syntology ran 3 of 3 samples · 0 unverified · 1 pointer-only (licence)We evaluate our instantiations at the scale of 125M to 1.
-
10 Jun 2024 3 repositories listed Syntology ran 4 of 4 samples · 0 unverified · 2 pointer-only (licence)Transformers with linear attention (i.
-
6 Jun 2024 3 repositories listedSpecifically, the Scaled Residual ConvMamba (SRCM) block is proposed to utilize the ability of Mamba to extract global features and convolution to enhance the local details to alleviate the issue that current…
-
15 Apr 2024 3 repositories listed Syntology ran 1 of 7 samples · 6 unverifiedThe release of nnU-Net marked a paradigm shift in 3D medical image segmentation, demonstrating that a properly configured U-Net architecture could still achieve state-of-the-art results.
-
11 Apr 2024 3 repositories listedThese spatial features then undergo intermediate temporal modeling facilitated by the Mamba block before progressing to the encoder section, which comprises vanilla upsampling Shift S-GCN blocks.
-
9 Apr 2024 3 repositories listed Syntology ran 3 of 3 samples · 0 unverified · 3 pointer-only (licence)Recent advancements in anomaly detection have seen the efficacy of CNN- and transformer-based approaches.
-
28 Mar 2024 3 repositories listed Syntology ran 8 of 10 samples · 2 unverifiedRemote sensing image classification forms the foundation of various understanding tasks, serving a crucial function in remote sensing image interpretation.
-
28 Mar 2024 3 repositories listedWe present Jamba, a new base large language model based on a novel hybrid Transformer-Mamba mixture-of-experts (MoE) architecture.
-
22 Mar 2024 3 repositories listed Syntology ran 4 of 4 samples · 0 unverified · 4 pointer-only (licence)Transformers have widely adopted attention networks for sequence mixing and MLPs for channel mixing, playing a pivotal role in achieving breakthroughs across domains.
-
28 Feb 2024 3 repositories listed Syntology ran 10 of 10 samples · 0 unverified · 1 pointer-only (licence)In this work, we explore whether we can improve language model efficiency (e.
-
1 Jul 2025 2 repositories listedNevertheless, neither the concept of hybrid Mamba and time-frequency attention models nor their generalization performance have been explored for speech enhancement.
-
1 Jun 2025 2 repositories listedVideo Frame Interpolation (VFI) is a fundamental Low-Level Vision (LLV) task that synthesizes intermediate frames between existing ones while maintaining spatial and temporal coherence.
-
23 May 2025 2 repositories listed Syntology ran 3 of 3 samples · 0 unverifiedStructured Linear Controlled Differential Equations (SLiCEs) provide a unifying framework for sequence models with structured, input-dependent state-transition matrices that retain the maximal expressivity of dense…
-
16 Feb 2025 2 repositories listedRecent 2D CNN-based domain adaptation approaches struggle with long-range dependencies due to limited receptive fields, making it difficult to adapt to target domains with significant spatial distribution changes.
-
14 Jan 2025 2 repositories listedTo address these, we propose a dual-domain hierarchical Mamba for MRI reconstruction from the following perspectives: (1) We pioneer vision Mamba in k-space learning.
-
23 Dec 2024 2 repositories listedAs such, conventional GNNs struggle to learn from these pathways due to the long-range dependencies of multiple pathways.
-
9 Dec 2024 2 repositories listed Syntology ran 0 of 13 samples · 13 unverifiedWe find that longer context models improve predictive performance -- our Mamba-based model surpasses the prior state-of-the-art on 9/14 tasks on the EHRSHOT prediction benchmark.
Syntology lines on 20 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections