Awesome-Vision-Mamba-Models
July 7, 2026 · View on GitHub
[NEWS.2024/11/10] The latest version of our paper (v3) is now available! This update includes numerous high-quality papers on visual Mamba.
[NEWS.2024/07/06] The updated version of our paper is now available!
[NEWS.2024/04/29] Our paper is released!
📢NOTE: If you have any questions, please don't hesitate to contact us at any of the following emails: xurui7943@gmail.com, syangcw@connect.ust.hk, ywangrm@connect.ust.hk, yu.cai@connect.ust.hk.
Mamba, a novel state space model, has gained recognition across diverse domains for its exceptional performance and efficient computational complexity. By addressing the limitations inherent in traditional visual foundation architectures, Mamba emerges as a promising contender poised to catalyze advancements in the field of computer vision.
:star: This repository hosts a curated collection of literature associated with Mamba models in computer vision. Feel free to star and fork. For further details, refer to the following paper:
Visual Mamba: A Survey and New Outlooks
Rui Xu, Shu Yang, Yihui Wang, Yu Cai, Bo Du, Hao Chen
SMART Lab, The Hong Kong University of Science and Technology
Contents
- Mamba
- Related Survey
- Visual Mamba Backbone Networks
- Vision Application (Modality)
- Valuable Insights
- Other Domains
Mamba
| Venue | Paper | Figure | Link | Code |
|---|---|---|---|---|
| COLM 2024 | Mamba: Linear-Time Sequence Modeling with Selective State Spaces | Link | Code | |
| ICML 2024 | Transformers are SSMs: Generalized Models and Efficient Algorithms Through Structured State Space Duality | Link | Code | |
| ICLR 2026 | Mamba-3: Improved Sequence Modeling using State Space Principles | Link | Code |
Related Survey
| Venue | Paper | Link |
|---|---|---|
| Applied Sciences 2024 | A Survey on Visual Mamba | Link |
| Engineering Applications of AI 2025 | Mamba-360: Survey of State Space Models as Transformer Alternative for Long Sequence Modelling: Methods, Applications, and Challenges | Link |
| IEEE TNNLS 2025 | Vision Mamba: A Comprehensive Survey and Taxonomy | Link |
| ACM TIST 2026 | A Survey of Mamba | Link |
Visual Mamba Backbone Networks
| Venue | Paper | Link | Code |
|---|---|---|---|
| ICML 2024 | Vision Mamba: Efficient Visual Representation Learning with Bidirectional State Space Model | Link | Code |
| NeurIPS 2024 | VMamba: Visual State Space Model | Link | Code |
| ECCV 2024 Oral | Mamba-ND: Selective State Space Modeling for Multi-Dimensional Data | Link | Code |
| ECCV 2024 Workshop | LocalMamba: Visual State Space Model with Windowed Selective Scan | Link | Code |
| AAAI 2025 | EfficientVMamba: Atrous Selective Scan for Light Weight Visual Mamba | Link | Code |
| BMVC 2024 | PlainMamba: Improving Non-Hierarchical Mamba in Visual Recognition | Link | Code |
| NeurIPS 2024 | Multi-Scale VMamba: Hierarchy in Hierarchy Visual State Space Model | Link | Code |
| NeurIPS 2024 | Vision Mamba Mender | Link | Code |
| NeurIPS 2024 | Exploring Token Pruning in Vision State Space Models | Link | |
| NeurIPS 2024 | QuadMamba: Learning Quadtree-based Selective Scan for Visual State Space Model | Link | Code |
| WACV 2025 | PTQ4VM: Post-Training Quantization for Visual Mamba | Link | Code |
| CVPR 2025 | Mamba-R: Vision Mamba Also Needs Registers | Link | Code |
| PRICAI 2025 | Vim-F: Visual State Space Model Benefiting from Learning in the Frequency Domain | Link | |
| ICLR 2025 | Autoregressive Pretraining with Mamba in Vision | Link | Code |
| CVPR 2025 | MambaVision: A Hybrid Mamba-Transformer Vision Backbone | Link | Code |
| CVPR 2025 | GroupMamba: Parameter-Efficient and Accurate Group Visual State Space Model | Link | Code |
| ICCV 2025 | VSSD: Vision Mamba with Non-Causal State Space Duality | Link | Code |
| ICML 2025 | Stochastic Layer-Wise Shuffle: A Stochastic Training Technique for Vision Mamba Models | Link | |
| AAAI 2025 | SparX: A Sparse Cross-Layer Connection Mechanism for Hierarchical Vision Mamba and Transformer Networks | Link | Code |
| ECCV 2024 Workshop | Famba-V: Fast Vision Mamba with Cross-Layer Token Fusion | Link | Code |
| IJCV 2026 | StableMamba: Distillation-free Scaling of Large SSMs for Images and Videos | Link | |
| CVPR 2025 | MAP: Unleashing Hybrid Mamba-Transformer Vision Backbone's Potential with Masked Autoregressive Pretraining | Link | Code |
| CVPR 2025 | Adventurer: Optimizing Vision Mamba Architecture Designs for Efficient Visual Recognition | Link | |
| ICLR 2025 | Spatial-Mamba: Effective Visual State Space Models via Structure-aware State Space Duality | Link | Code |
| CVPR 2025 | EfficientViM: Efficient Vision Mamba with Hidden State Mixer based State Space Duality | Link | Code |
| CVPR 2025 | MobileMamba: Lightweight Multi-Receptive Visual Mamba Network | Link | Code |
| ICCV 2025 | TinyViM: Frequency Decoupling for Tiny Hybrid Vision Mamba | Link | |
| CVPR 2025 | GG-SSMs: Graph-Generating State Space Models | Link | |
| ICLR 2025 | MambaQuant: Quantizing the Mamba Family with Variance Aligned Rotation Methods | Link | Code |
| CVPR 2025 | DefMamba: Deformable Visual State Space Model | Link | |
| CVPR 2025 | Mamba-Adaptor: State Space Model Adaptor for Visual Recognition | Link | |
| ICCV 2025 | PVMamba: Parallelizing Vision Mamba via Dynamic State Aggregation | Link | |
| NeurIPS 2025 | DAMamba: Vision State Space Model with Dynamic Adaptive Scan | Link | Code |
| NeurIPS 2025 | TF-MAS: Training-free Mamba2 Architecture Search | Link | |
| ICCV 2025 | OuroMamba: A Data-Free Quantization Framework for Vision Mamba | Link | |
| ICASSP 2025 | MambaNext: An Enhanced Backbone Network with Focus Linear Attention | Link | |
| AAAI 2026 | 2D-CrossScan Mamba: Enhancing State Space Models with Spatially Consistent Multi-Path 2D Information Propagation | Link | |
| WACV 2026 | Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing | Link | Code |
| CVPR 2026 | HAMSA: Scanning-Free Vision State Space Models via SpectralPulseNet | Link | |
| ICML 2026 | Partial Ring Scan: Revisiting Scan Order in Vision State Space Models | Link | |
| ICML 2026 | Deformba: Vision State Space Model with Adaptive State Fusion | Link | Code |
| ICML 2026 | Spatial-Aware Reduction Framework: Towards Efficient and Faithful Visual State Space Models | Link | |
| ICML 2026 | SF-Mamba: Rethinking State Space Model for Vision | Link | |
| ICLR 2026 | Enabling True Global Perception in State Space Models for Visual Tasks | Link |
Vision Application
Image
Natural Image
| Venue | Paper | Link | Code | Task |
|---|---|---|---|---|
| ECCV 2024 | MambaIR: A Simple Baseline for Image Restoration with State-Space Model | Link | Code | Image Restoration |
| IEEE TCSVT 2025 | VmambaIR: Visual State Space Model for Image Restoration | Link | Code | Image Restoration |
| ECCV 2024 | ZigMa: A DiT-style Zigzag Mamba Diffusion Model | Link | Code | Generation |
| ACM MM 2024 | Learning Enriched Features via Selective State Spaces Model for Efficient Image Deblurring | Link | Deblurring | |
| IEEE TPAMI 2025 | Gamba: Marry Gaussian Splatting with Mamba for single view 3D reconstruction | Link | 3D Reconstruction | |
| NeurIPS 2024 | MambaAD: Exploring State Space Models for Multi-class Unsupervised Anomaly Detection | Link | code | Anomaly Detection |
| ACM MM 2024 | DGMamba: Domain Generalization via Generalized State Space Model | Link | Code | Domain Generalization |
| ACM MM 2024 | FreqMamba: Viewing Mamba from a Frequency Perspective for Image Deraining | Link | Code | Deraining |
| MIPR 2024 | CU-Mamba: Selective State Space Models with Channel Learning for Image Restoration | Link | Image Restoration | |
| CVPR 2024 Workshop | DVMSR: Distillated Vision Mamba for Efficient Super-Resolution | Link | Code | Super-Resolution |
| ICONIP 2024 | Retinexmamba: Retinex-based Mamba for Low-light Image Enhancement | Link | Code | Image Enhancement |
| NeurIPS 2024 | MambaLLIE: Implicit Retinex-Aware Low Light Enhancement with Global-then-Local State Space | Link | Code | Image Enhancement |
| WACV 2025 | SUM: Saliency Unification through Mamba for Visual Attention Modeling | Link | Code | Saliency Prediction |
| ECCV 2024 | MTMamba: Enhancing Multi-Task Dense Scene Understanding by Mamba-Based Decoders | Link | Code | Multi-Task Dense Prediction |
| ICML 2024 Workshop | Parallelizing Autoregressive Generation with Variational State Space Models | Link | Generation | |
| PRCV 2024 | ALMRR: Anomaly Localization Mamba on Industrial Textured Surface with Feature Reconstruction and Refinement | Link | Code | Anomaly Detection |
| NeurIPS 2024 | Hamba: Single-view 3D Hand Reconstruction with Graph-guided Bi-Scanning Mamba | Link | Code | 3D Hand Mesh Recovery |
| BMVC 2024 | MxT: Mamba x Transformer for Image Inpainting | Link | Image Inpainting | |
| WBIR 2024 Workshop | Mamba? Catch The Hype Or Rethink What Really Helps for Image Registration | Link | Code | Image Registration |
| ACM MM 2024 | Wave-Mamba: Wavelet State Space Model for Ultra-High-Definition Low-Light Image Enhancement | Link | Code | Image Enhancement |
| AAAI 2025 | ZeroMamba: Exploring Visual State Space Model for Zero-Shot Learning | Link | Code | Zero-Shot Learning |
| ICPR 2024 | DS MYOLO: A Reliable Object Detector Based on SSMs for Driving Scenarios | Link | Object Detection | |
| IEEE TITS 2025 | DSDFormer: An Innovative Transformer-Mamba Framework for Robust High-Precision Driver Distraction Identification | Link | Image Classification | |
| IEEE TCSVT 2025 | Retinex-RAWMamba: Bridging Demosaicing and Denoising for Low-Light RAW Image Enhancement | Link | Code | Image Enhancement |
| WACV 2025 | Mamba-ST: State Space Model for Efficient Style Transfer | Link | Code | Style Transfer |
| ACCV 2024 | OneBEV: Using One Panoramic Image for Bird's-Eye-View Semantic Mapping | Link | Code | Bird's-Eye-View Semantic Mapping |
| Neurocomputing 2024 | MambaTSR: You only need 90k parameters for traffic sign recognition | Link | Code | Image Classification |
| Scientific Reports 2024 | Toward identity preserving in face sketch-photo synthesis using a hybrid CNN-Mamba framework | Link | Image Synthesis | |
| NeurIPS 2024 | Hybrid Mamba for Few-Shot Segmentation | Link | Code | Few-Shot Segmentation |
| NeurIPS 2024 | START: A Generalized State Space Model with Saliency-Driven Token-Aware Transformation | Link | Code | Domain Generalization |
| IJCV 2025 | Mamba Capsule Routing Towards Part-Whole Relational Camouflaged Object Detection | Link | Code | Camouflaged Object Detection |
| Automation in Construction 2024 | Topology-aware Mamba for Crack Segmentation in Structures | Link | Code | Crack Segmentation |
| ACCV 2024 | Wavelet-based Mamba with Fourier Adjustment for Low-light Image Enhancement | Link | Code | Image Enhancement |
| Image and Vision Computing 2025 | ShadowMamba: State-Space Model with Boundary-Region Selective Scan for Shadow Removal | Link | Shadow Removal | |
| NeurIPS 2024 | ECMamba: Consolidating Selective State Space Model with Retinex Guidance for Efficient Multiple Exposure Correction | Link | Code | Exposure Correction |
| NeurIPS 2024 | DiMSUM: Diffusion Mamba -- A Scalable and Unified Spatial-Frequency Method for Image Generation | Link | Code | Generation |
| ACM MM 2024 | Realistic Full-Body Motion Generation from Sparse Tracking with State Space Model | Link | Motion Generation | |
| WACV 2025 | SEM-Net: Efficient Pixel Modelling for Image Inpainting with Spatially Enhanced SSM | Link | Code | Image Inpainting |
| IEEE SPL 2024 | LFSamba: Marry SAM with Mamba for Light Field Salient Object Detection | Link | Code | Salient Object Detection |
| NeurIPS 2024 Workshop | Diverse capability and scaling of diffusion and auto-regressive models when learning abstract rules | Link | Rule Learning/Reasoning | |
| CVPR 2025 | OSMamba: Omnidirectional Spectral Mamba with Dual-Domain Prior Generator for Exposure Correction | Link | Exposure Correction | |
| AAAI 2025 | Selective Visual Prompting in Vision Mamba | Link | Code | Domain Adaptation |
| IEEE TII 2025 | VarAD: Lightweight High-Resolution Image Anomaly Detection via Visual Autoregressive Modeling | Link | Code | Anomaly Detection |
| CVPR 2025 | Efficient Visual State Space Model for Image Deblurring | Link | Code | Deblurring |
| ACCV 2024 | Image Deraining with Frequency-Enhanced State Space Model | Link | Deraining | |
| AAAI 2025 | Mamba YOLO: A Simple Baseline for Object Detection with State Space Model | Link | Code | Object Detection |
| ICML 2025 | FourierMamba: Fourier Learning Integration with State Space Models for Image Deraining | Link | Deraining | |
| ICML 2025 | QMamba: On First Exploration of Vision Mamba for Image Quality Assessment | Link | Image Quality Assessment | |
| ACCV 2024 | Mamba-based Light Field Super-Resolution with Efficient Subspace Scanning | Link | Super-Resolution | |
| ACCV 2024 | PixMamba: Leveraging State Space Models in a Dual-Level Architecture for Underwater Image Enhancement | Link | Code | Image Enhancement |
| AAAI 2025 | Pose Magic: Efficient and Temporally Consistent Human Pose Estimation with a Hybrid Mamba-GCN Network | Link | 3D Human Pose Estimation | |
| AAAI 2025 | PoseMamba: Monocular 3D Human Pose Estimation with Bidirectional Global-Local Spatio-Temporal State Space Model | Link | 3D Human Pose Estimation | |
| IEEE TIFS 2026 | Neural Architecture Search-Based Global–Local Vision Mamba for Palm-Vein Recognition | Link | Palm-Vein Recognition | |
| CVPR 2025 | QMambaBSR: Burst Image Super-Resolution with Query State Space Model | Link | Code | Super-Resolution |
| IEEE TPAMI 2025 | MTMamba++: Enhancing Multi-Task Dense Scene Understanding via Mamba-Based Decoders | Link | Multi-Task Dense Prediction | |
| ICLR 2025 | MambaPEFT: Exploring Parameter-Efficient Fine-Tuning for Mamba | Link | Parameter-Efficient Fine-Tuning | |
| CVPR 2025 | Parameter Efficient Mamba Tuning via Projector-targeted Diagonal-centric Linear Transformation | Link | Parameter-Efficient Fine-Tuning | |
| CVPR 2025 | MambaIRv2: Attentive State Space Restoration | Link | Code | Image Restoration |
| CVPR 2025 Workshop | XYScanNet: An Interpretable State Space Model for Perceptual Image Deblurring | Link | Deblurring | |
| ICASSP 2025 | ViM-Disparity: Bridging the Gap of Speed, Accuracy and Memory for Disparity Map Generation | Link | Code | Depth Estimation |
| CVPR 2025 | MaIR: A Locality- and Continuity-Preserving Mamba for Image Restoration | Link | Code | Image Restoration |
| WACV 2025 | EDMB: Edge Detector with Mamba | Link | Code | Edge Detection |
| ACM MM 2025 | WMamba: Wavelet-based Mamba for Face Forgery Detection | Link | Forgery Detection | |
| IEEE TIP 2025 | UniUIR: Considering Underwater Image Restoration as An All-in-One Learner | Link | Image Restoration | |
| CVPR 2025 Highlight | Mamba as a Bridge: Where VFMs Meet VLMs for Domain-Generalized Semantic Segmentation | Link | Code | Semantic Segmentation |
| CVPR 2025 | Mesh Mamba: A Unified State Space Model for Saliency Prediction | Link | Saliency Prediction | |
| CVPR 2025 | MambaIC: State Space Models for High-Performance Learned Image Compression | Link | Code | Image Compression |
| CVPR 2025 | MambaFlow: A Mamba-Centric Architecture for End-to-End Optical Flow Estimation | Link | Optical Flow Estimation | |
| CVPR 2025 | JamMa: Ultra-lightweight Local Feature Matching with Joint Mamba | Link | Feature Matching | |
| CVPR 2025 | Enhancing Online Continual Learning with Plug-and-Play State Space Model | Link | Continual Learning | |
| ICCV 2025 | MambaML: Exploring State Space Models for Multi-Label Image Classification | Link | Multi-Label Classification | |
| ICCV 2025 | Cassic: Towards Content-Adaptive State-Space Models for Learned Image Compression | Link | Image Compression | |
| ICCV 2025 | EAMamba: Efficient All-Around Vision State Space Model for Image Restoration | Link | Image Restoration | |
| ICCV 2025 | MeshMamba: State Space Models for Articulated 3D Mesh Generation and Reconstruction | Link | 3D Human Mesh Recovery | |
| IJCAI 2025 | Directing Mamba to Complex Textures: An Efficient Texture-Aware State Space Model for Image Restoration | Link | Image Restoration | |
| ACM MM 2025 | DeflareMamba: Hierarchical Vision Mamba for Contextually Consistent Lens Flare Removal | Link | Code | Image Restoration |
| ACM MM 2025 | UIS-Mamba: Exploring Mamba for Underwater Instance Segmentation | Link | Code | Instance Segmentation |
| AAAI 2025 | SalM²: An Extremely Lightweight Saliency Mamba Model for Real-Time Cognitive Awareness of Driver Attention | Link | Saliency Prediction | |
| CVPR 2025 | SCSegamba: Lightweight Structure-Aware Vision Mamba for Crack Segmentation in Structures | Link | Crack Segmentation | |
| CVPR 2025 | Binarized Mamba-Transformer for Lightweight Quad Bayer HybridEVS Demosaicing | Link | Image Demosaicing | |
| IJCAI 2025 | Omni-Dimensional State Space Model-driven SAM for Pixel-level Anomaly Detection | Link | Anomaly Detection | |
| ACM MM 2025 | ACMamba: Fast Unsupervised Anomaly Detection via An Asymmetrical Consensus State Space Model | Link | Anomaly Detection | |
| IEEE Transactions on Cybernetics 2025 | Parameter Aware Mamba Model for Multi-task Dense Prediction | Link | Multi-Task Dense Prediction | |
| ICASSP 2025 | ICAA-Mamba: Vision Mamba for Image Color Aesthetics Assessment | Link | Image Quality Assessment | |
| ICASSP 2025 | Cross-Modality Fusion Mamba for All-in-One Extreme Weather-Degraded Image Restoration | Link | Image Restoration | |
| ICASSP 2025 | EPI-Mamba: State Space Model for Semantic Segmentation from Light Fields | Link | Semantic Segmentation | |
| ICASSP 2025 | MambaInst: Lightweight State Space Model for Real-Time Instance Segmentation | Link | Instance Segmentation | |
| ICASSP 2025 | Vision Mamba-Based Approach for Incomplete Boundary Document Image Rectification | Link | Image Restoration | |
| ICASSP 2025 | HandS3C: 3D Hand Mesh Reconstruction with State Space Spatial Channel Attention from RGB Images | Link | 3D Hand Mesh Recovery | |
| ICASSP 2025 | InsectMamba: State Space Model with Adaptive Composite Features for Insect Recognition | Link | Image Classification | |
| ICASSP 2025 | StereoMamba: Enhancing Stereo Image Super-Resolution with Structured State Space Models and Bi-Directional Cross Attention | Link | Super-Resolution | |
| ICASSP 2025 | MS-RainMamba: Learning Multi-Scale State Space Models for Single Image Deraining | Link | Deraining | |
| ICASSP 2025 | RestorMamba: An Enhanced Synergistic State Space Model for Image Restoration | Link | Image Restoration | |
| ICASSP 2025 Oral | First-order State Space Model for Lightweight Image Super-resolution | Link | Super-Resolution | |
| ECAI 2025 | DA-Mamba: Domain Adaptive Hybrid Mamba-Transformer Based One-Stage Object Detection | Link | Code | Object Detection |
| AAAI 2026 | Depth-Synergized Mamba Meets Memory Experts for All-Day Image Reflection Separation | Link | Image Restoration | |
| AAAI 2026 | Rectification Reimagined: A Unified Mamba Model for Image Correction and Rectangling with Prompts | Link | Image Correction | |
| AAAI 2026 | Disentangled Hypergraph-Guided Mamba Scanning for Fine-Grained Visual Recognition | Link | Image Classification | |
| ICLR 2026 | WIMFRIS: WIndow Mamba Fusion and Parameter Efficient Tuning for Referring Image Segmentation | Link | Referring Image Segmentation | |
| ICLR 2026 | Content-Aware Mamba for Learned Image Compression | Link | Image Compression | |
| WACV 2026 | DF-Mamba: Deformable State Space Modeling for 3D Hand Pose Estimation in Interactions | Link | 3D Hand Pose Estimation | |
| WACV 2026 | SasMamba: A Lightweight Structure-Aware Stride State Space Model for 3D Human Pose Estimation | Link | 3D Human Pose Estimation | |
| WACV 2026 | Forensim: Can Image Splicing and Copy-Move Forgery Be Detected by the Same Model? An Attention-Based State-Space Approach | Link | Forgery Detection | |
| WACV 2026 | D2Mamba: Dual Domain Guided Informed Search in State Space Model for Underwater Image Enhancement | Link | Image Enhancement | |
| WACV 2026 | From Darkness to Detail: Frequency-Aware SSMs for Low-Light Vision | Link | Code | Image Enhancement |
| WACV 2026 | Codebook Knowledge with Mamba-Transformer For Low-Light Image Enhancement | Link | Image Enhancement | |
| CVPR 2026 | DA-Mamba: Learning Domain-Aware State Space Model for Global-Local Alignment in Domain Adaptive Object Detection | Link | Object Detection | |
| CVPR 2026 | MambaSIC: Mamba-based Stereo Image Compression with Bi-directional Multi-reference Entropy Model | Link | Image Compression | |
| CVPR 2026 | MambaCS: Multi-Scale Gradient-Guided Unrolling Architecture with Adaptive Mamba for Compressive Sensing | Link | Code | Compressive Sensing |
| CVPR 2026 | MixerCSeg: An Efficient Mixer Architecture for Crack Segmentation via Decoupled Mamba Attention | Link | Code | Crack Segmentation |
| CVPR 2026 | CrackSSM: Reviving SSMs for Crack Segmentation via Dynamic Scanning | Link | Crack Segmentation | |
| CVPR 2026 | AKCMamba-YOLO: Selective State Space Models For Real-Time Object Detection | Link | Object Detection | |
| CVPR 2026 | SSM-Aware Token-Efficient VMamba via Adaptive Patch Pruning and Merging for Person Re-Identification | Link | Person Re-Identification | |
| CVPR 2026 | Scalable Feature Matching via State Space Modeling and Sparse Correlation | Link | Code | Feature Matching |
| CVPR 2026 Findings | Q-MambaIR: Accurate Quantized Mamba for Efficient Image Restoration | Link | Image Restoration | |
| ICASSP 2026 | STYMAM: A Mamba-Based Generator for Artistic Style Transfer | Link | Style Transfer | |
| ICASSP 2026 | Light Field Image Super-Resolution with Multi-Scale Context Aggregation Mamba | Link | Super-Resolution | |
| ECCV 2026 | MambaRaw: Selective State Space Modeling for Efficient 4K Raw Image Reconstruction | Link | Code | Image Reconstruction |
Remote Sensing Image
| Venue | Paper | Link | Code | Task |
|---|---|---|---|---|
| IEEE TGRS 2024 | MiM-ISTD: Mamba-in-Mamba for Efficient Infrared Small Target Detection | Link | Code | Object Detection |
| IEEE GRSL 2024 | RSMamba: Remote Sensing Image Classification with State Space Model | Link | Code | Remote Sensing Image Classification |
| Heliyon 2024 | Samba: Semantic Segmentation of Remotely Sensed Images with State Space Model | Link | Code | Semantic Segmentation |
| IEEE GRSL 2024 | RS3Mamba: Visual State Space Model for Remote Sensing Images Semantic Segmentation | Link | Code | Semantic Segmentation |
| IEEE TGRS 2025 | RS-Mamba for Large Remote Sensing Image Dense Prediction | Link | Code | Semantic Segmentation |
| IEEE TGRS 2024 | ChangeMamba: Remote Sensing Change Detection with Spatio-Temporal State Space Model | Link | Code | Change Detection |
| IEEE TGRS 2024 | SSUMamba: Spatial-Spectral Selective State Space Model for Hyperspectral Image Denoising | Link | Code | Hyperspectral Image Denoising |
| IEEE TMM 2024 | Frequency-Assisted Mamba for Remote Sensing Image Super-Resolution | Link | Code | Super-Resolution |
| IEEE TGRS 2025 | GraphMamba: An Efficient Graph Structure Learning Vision Mamba for Hyperspectral Image Classification | Link | Code | Hyperspectral Image Classification |
| IEEE TGRS 2025 | DMM: Disparity-guided Multispectral Mamba for Oriented Object Detection in Remote Sensing | Link | Code | Oriented Object Detection |
| IEEE TGRS 2025 | HTD-Mamba: Efficient Hyperspectral Target Detection with Pyramid State Space Model | Link | Code | Hyperspectral Target Detection |
| IEEE TGRS 2025 | DualMamba: A Lightweight Spectral-Spatial Mamba-Convolution Network for Hyperspectral Image Classification | Link | Hyperspectral Image Classification | |
| IEEE TGRS 2025 | CDMamba: Incorporating Local Clues Into Mamba for Remote Sensing Image Binary Change Detection | Link | Code | Change Detection |
| IEEE TGRS 2025 | 3DSS-Mamba: 3D-Spectral-Spatial Mamba for Hyperspectral Image Classification | Link | Hyperspectral Image Classification | |
| Neurocomputing 2025 | Mamba-in-Mamba: Centralized Mamba-Cross-Scan in Tokenized Mamba Model for Hyperspectral Image Classification | Link | Code | Hyperspectral Image Classification |
| IEEE JSTARS 2024 | Rethinking Scanning Strategies with Vision Mamba in Semantic Segmentation of Remote Sensing Imagery: An Experimental Study | Link | Semantic Segmentation | |
| IEEE TGRS 2024 | MambaHSI: Spatial–Spectral Mamba for Hyperspectral Image Classification | Link | Code | Hyperspectral Image Classification |
| Remote Sensing Letters 2025 | Multi-head Spatial-Spectral Mamba for Hyperspectral Image Classification | Link | Code | Hyperspectral Image Classification |
| IEEE GRSL 2024 | WaveMamba: Spatial-Spectral Wavelet Mamba for Hyperspectral Image Classification | Link | Code | Hyperspectral Image Classification |
| Neurocomputing 2025 | Spatial-Spectral Morphological Mamba for Hyperspectral Image Classification | Link | Code | Hyperspectral Image Classification |
| IEEE GRSL 2024 | UNetMamba: An Efficient UNet-Like Mamba for Semantic Segmentation of High-Resolution Remote Sensing Images | Link | Code | Semantic Segmentation |
| IEEE GRSL 2024 | MambaFormerSR: A Lightweight model for Remote-Sensing Image Super-Resolution | Link | Super-Resolution | |
| Scientific Reports 2024 | YOLOv5_mamba: unmanned aerial vehicle object detection based on bidirectional dense feedback network and adaptive gate feature fusion | Link | Code | Object Detection |
| ECML/PKDD 2024 Workshop | A Deep Learning-Based Approach for Mangrove Monitoring | Link | Code | Semantic Segmentation |
| IEEE TGRS 2024 | HyperMamba: A Spectral-Spatial Adaptive Mamba for Hyperspectral Image Classification | Link | Code | Hyperspectral Image Classification |
| ACM MM 2024 | VmambaSCI: Dynamic Deep Unfolding Network with Mamba for Compressive Spectral Imaging | Link | Spectral Compressive Imaging | |
| IEEE TGRS 2024 | ConMamba: CNN and SSM High-Performance Hybrid Network for Remote Sensing Change Detection | Link | Change Detection | |
| IEEE TGRS 2024 | A Novel Remote Sensing Image Change Detection Approach Based on Multi-level State Space Model | Link | Code | Change Detection |
| IEEE TGRS 2024 | Dynamic Token Augmentation Mamba for Cross-Scene Classification of Hyperspectral Image | Link | Code | Hyperspectral Image Classification |
| IEEE GRSL 2024 | PPMamba:Enhancing Semantic Segmentation in Remote Sensing Imagery by SS2D | Link | Code | Semantic Segmentation |
| AAAI 2025 | Detail Matters: Mamba-Inspired Joint Unfolding Network for Snapshot Spectral Compressive Imaging | Link | Code | Spectral Compressive Imaging |
| IGARSS 2025 | Mamba-MOC: A Multicategory Remote Object Counting via State Space Model | Link | Code | Object Counting |
| IEEE TGRS 2025 | S2Mamba: A Spatial-spectral State Space Model for Hyperspectral Image Classification | Link | Hyperspectral Image Classification | |
| Remote Sensing 2024 | Spectral-Spatial Mamba for Hyperspectral Image Classification | Link | Hyperspectral Image Classification | |
| WACV 2025 | A Mamba-based Siamese Network for Remote Sensing Change Detection | Link | Change Detection | |
| IEEE TGRS 2025 | MSFMamba: Multi-Scale Feature Fusion State Space Model for Multi-Source Remote Sensing Image Classification | Link | Remote Sensing Image Classification | |
| IEEE TGRS 2025 | IGroupSS-Mamba: Interval Group Spatial–Spectral Mamba for Hyperspectral Image Classification | Link | Hyperspectral Image Classification | |
| IGARSS 2025 | SITSMamba for Crop Classification based on Satellite Image Time Series | Link | Code | Remote Sensing Image Classification |
| ICASSP 2025 | UV-Mamba: A DCN-Enhanced State Space Model for Urban Village Boundary Identification in High-Resolution Remote Sensing Images | Link | Semantic Segmentation | |
| ICASSP 2026 | RemoteDet-Mamba: A Hybrid Mamba-CNN Network for Multi-modal Object Detection in Remote Sensing Images | Link | Object Detection | |
| IGARSS 2025 | WSSM: Geographic-enhanced hierarchical state-space model for global station weather forecast | Link | Weather Forecasting | |
| IEEE GRSL 2025 | CDxLSTM: Boosting Remote Sensing Change Detection With Extended Long Short-Term Memory | Link | Code | Change Detection |
| IEEE TGRS 2025 | IRSRMamba: Infrared Image Super-Resolution via Mamba-based Wavelet Transform Feature Modulation Model | Link | Code | Super-Resolution |
| AAAI 2025 | DehazeMamba: SAR-guided Optical Remote Sensing Image Dehazing with Adaptive State Space | Link | Code | Dehazing |
| IJCAI 2025 | HSRMamba: Contextual Spatial-Spectral State Space Model for Single Hyperspectral Image Super-Resolution | Link | Code | Super-Resolution |
| NeurIPS 2025 | RoMA: Scaling up Mamba-based Foundation Models for Remote Sensing | Link | Remote Sensing Foundation Model | |
| IEEE TGRS 2025 | Wavelet-Assisted Mamba for Satellite-Derived Sea Surface Temperature Super-Resolution | Link | Super-Resolution | |
| IJCAI 2025 | VimGeo: Efficient Cross-View Geo-Localization with Vision Mamba Architecture | Link | Code | Geo-Localization |
| IJCAI 2025 | DPMamba: Distillation Prompt Mamba for Multimodal Remote Sensing Image Classification with Missing Modalities | Link | Remote Sensing Image Classification | |
| ICASSP 2025 | SSRMamba: Efficient Visual State Space Model for Spectral Super-Resolution | Link | Super-Resolution | |
| ICASSP 2025 | SSFMamba: Spatial-Spectral Fusion State Space Model for Pansharpening | Link | Pansharpening | |
| AAAI 2026 | MFmamba: A Multi-function Network for Panchromatic Image Resolution Restoration Based on State-Space Model | Link | Pansharpening | |
| AAAI 2026 | M3SR: Multi-Scale Multi-Perceptual Mamba for Efficient Spectral Reconstruction | Link | Hyperspectral Reconstruction | |
| AAAI 2026 | MMMamba: A Versatile Cross-Modal in Context Fusion Framework for Pan-Sharpening and Zero-Shot Image Enhancement | Link | Pansharpening | |
| WACV 2026 | DMS2F-HAD: A Dual-branch Mamba-based Spatial-Spectral Fusion Network for Hyperspectral Anomaly Detection | Link | Code | Hyperspectral Anomaly Detection |
| ICASSP 2026 | SAR Ship Wake Detection Based on Siamese Network with Mamba Cross-Domain Feature Fusion | Link | Object Detection |
Medical Image
| Venue | Paper | Link | Code | Task |
|---|---|---|---|---|
| MICCAI 2024 | SegMamba: Long-range Sequential Modeling Mamba For 3D Medical Image Segmentation | Link | Code | 3D Medical Segmentation |
| ISBI 2025 | nnMamba: 3D Biomedical Image Segmentation, Classification and Landmark Detection with State Space Model | Link | Code | 3D Medical Segmentation |
| ACM TOMCCAP 2025 | VM-UNet: Vision Mamba UNet for Medical Image Segmentation | Link | Code | 2D Medical Segmentation |
| MICCAI 2024 | Swin-UMamba: Mamba-based UNet with ImageNet-based pretraining | Link | Code | 2D Medical Segmentation |
| KBS 2024 | Semi-Mamba-UNet: Pixel-Level Contrastive Cross-Supervised Visual Mamba-based UNet for Semi-Supervised Medical Image Segmentation | Link | Code | 2D Medical Segmentation |
| BIBM 2024 | MamMIL: Multiple Instance Learning for Whole Slide Images with State Space Models | Link | Cancer Subtyping | |
| MICCAI 2024 | MambaMIL: Enhancing Long Sequence Modeling with Sequence Reordering in Computational Pathology | Link | Code | Cancer Subtyping/Survival Prediction |
| MICCAI 2024 | LKM-UNet: Large Kernel Vision Mamba UNet for Medical Image Segmentation | Link | Code | 2D Medical Segmentation |
| BIBM 2024 | MD-Dose: A diffusion model based on the Mamba for radiation dose prediction | Link | Code | Radiation Dose Prediction |
| ISBRA 2024 | VM-UNET-V2 Rethinking Vision Mamba UNet for Medical Image Segmentation | Link | Code | 2D Medical Segmentation |
| Neurocomputing 2025 | H-vmunet: High-order Vision Mamba UNet for Medical Image Segmentation | Link | Code | 2D Medical Segmentation |
| MIDL 2024 | ViM-UNet: Vision Mamba for Biomedical Segmentation | Link | Code | 2D Medical Segmentation |
| MICCAI 2024 | nnU-Net Revisited: A Call for Rigorous Validation in 3D Medical Image Segmentation | Link | Code | 3D Medical Segmentation |
| CVPR 2024 Workshop | Vim4Path: Self-Supervised Vision Mamba for Histopathology Images | Link | Code | Cancer Subtyping |
| MIPR 2024 | UU-Mamba: Uncertainty-aware U-Mamba for Cardiac Image Segmentation | Link | 3D Medical Segmentation | |
| MICCAI 2024 Oral | Cardiovascular Disease Detection from Multi-View Chest X-rays with BI-Mamba | Link | Code | Risk Prediction |
| Scientific Reports 2025 | Combining Graph Neural Network and Mamba to Capture Local and Global Tissue Spatial Relationships in Whole Slide Images | Link | Code | Cancer Subtyping/Survival Prediction |
| WACV 2025 | Convolution and Attention-Free Mamba-based Cardiac Image Segmentation | Link | Code | 2D Medical Segmentation |
| BMVC 2024 | On Evaluating Adversarial Robustness of Volumetric Medical Segmentation Models | Link | Code | 3D Medical Segmentation |
| MICCAI 2024 Workshop | Vision Mamba for Classification of Breast Ultrasound Images | Link | Medical Image Classification | |
| MICCAI 2024 | Deform-Mamba Network for MRI Super-Resolution | Link | Super-Resolution | |
| ICPR 2024 | Self-Prior Guided Mamba-UNet Networks for Medical Image Super-Resolution | Link | Super-Resolution | |
| KDD Workshop 2024 | State Space Model-based Classification of Major Depressive Disorder Across Multiple Imaging Sites | Link | Medical Image Classification | |
| MICCAI 2024 | ShapeMamba-EM: Fine-Tuning Foundation Model with Local Shape Descriptors and Mamba Blocks for 3D EM Image Segmentation | Link | 3D Medical Segmentation | |
| ICME 2025 | MambaMIC: An Efficient Baseline for Microscopic Image Classification with State Space Models | Link | Code | Medical Image Classification |
| ICASSP 2025 | SX-Stitch: An Efficient VMS-UNet Based Framework for Intraoperative Scoliosis X-Ray Image Stitching | Link | Medical Image Stitching | |
| ICASSP 2025 | MpoxMamba: A Grouped Mamba-based Lightweight Hybrid Network for Mpox Detection | Link | Code | Medical Image Classification |
| Scientific Reports 2024 | A mixed Mamba U-net for prostate segmentation in MR images | Link | 3D Medical Segmentation | |
| IEEE TMI 2025 | Serp-Mamba: Advancing High-Resolution Retinal Vessel Segmentation with Selective State-Space Model | Link | 2D Medical Segmentation | |
| MICCAI 2024 | Tri-Plane Mamba: Efficiently Adapting Segment Anything Model for 3D Medical Images | Link | Code | 3D Medical Segmentation |
| ACCV 2024 Workshop | SkinMamba: A Precision Skin Lesion Segmentation Architecture with Cross-Scale Global State Modeling and Frequency Boundary Guidance | Link | Code | 2D Medical Segmentation |
| IEEE Sensors Journal 2025 | SPRMamba: Surgical Phase Recognition for Endoscopic Submucosal Dissection with Mamba | Link | Surgical Phase Recognition | |
| WACV 2025 | MambaRecon: MRI Reconstruction with Structured State Space Models | Link | Code | Image Reconstruction |
| MICCAI 2024 | EM-Net: Efficient Channel and Frequency Learning with Mamba for 3D Medical Image Segmentation | Link | Code | 3D Medical Segmentation |
| MICCAI 2024 | MetaUNETR: Rethinking Token Mixer Encoding for Efficient Multi-organ Segmentation | Link | Code | 3D Medical Segmentation |
| MICCAI 2024 | PathMamba: Weakly Supervised State Space Model for Multi-class Segmentation of Pathology Images | Link | Code | 2D Medical Segmentation |
| MICCAI 2024 | Efficient and Gender-adaptive Graph Vision Mamba for Pediatric Bone Age Assessment | Link | Code | Bone Age Assessment |
| MICCAI 2024 | Polyp-Mamba: Polyp Segmentation with Visual Mamba | Link | Polyp Segmentation | |
| IEEE TMI 2024 | Unleash the Power of State Space Model for Whole Slide Image with Local Aware Scanning and Importance Resampling | Link | Code | Cancer Subtyping/Survival Prediction |
| IEEE TMI 2024 | Swin-UMamba+: Adapting Mamba-based vision foundation models for medical image segmentation | Link | Code | 2D & 3D Medical Segmentation |
| AAAI 2025 | S3Mamba: Small-Size-Sensitive Mamba for Lesion Segmentation | Link | Code | 2D Medical Segmentation |
| ICME 2025 | HCMA-UNet: A Hybrid CNN-Mamba UNet with Axial Self-Attention for Efficient Breast Cancer Segmentation | Link | Code | 3D Medical Segmentation |
| IEEE TMI 2025 | Merging Context Clustering with Visual State Space Models for Medical Image Segmentation | Link | Code | 2D Medical Segmentation |
| ISBI 2025 | GLFC: Unified Global-Local Feature and Contrast Learning with Mamba-Enhanced UNet for Synthetic CT Generation from CBCT | Link | Code | Image Synthesis |
| IEEE TCSVT 2025 | DH-Mamba: Exploring Dual-Domain Hierarchical State Space Models for MRI Reconstruction | Link | Code | Image Reconstruction |
| IEEE TCSS 2025 | MSV-Mamba: A Multiscale Vision Mamba Network for Echocardiography Segmentation | Link | 2D Medical Segmentation | |
| Information Fusion 2025 | Polyp-Mamba: A Hybrid Multi-Frequency Perception Gated Selection Network for polyp segmentation | Link | Polyp Segmentation | |
| Patterns 2025 | UltraLight VM-UNet: Parallel Vision Mamba Significantly Reduces Parameters for Skin Lesion Segmentation | Link | Code | 2D Medical Segmentation |
| IEEE TMM 2026 | T-Mamba: A Unified Framework with Long-Range Dependency in Dual-Domain for 2D & 3D Tooth Segmentation | Link | Code | 2D & 3D Medical Segmentation |
| MICCAI 2025 | Sparse Reconstruction of Optical Doppler Tomography with Alternative State Space Model | Link | Image Reconstruction | |
| Exploration of Medicine 2025 | MUCM-Net: A Mamba Powered UCM-Net for Skin Lesion Segmentation | Link | Code | 2D Medical Segmentation |
| ICCV 2025 | TokenUnify: Scaling Up Autoregressive Pretraining for Neuron Segmentation | Link | 3D Medical Segmentation | |
| IEEE JBHI 2025 | SliceMamba with Neural Architecture Search for Medical Image Segmentation | Link | 2D Medical Segmentation | |
| MIA 2025 | MambaMIM: Pre-training Mamba with State Space Token Interpolation | Link | Code | Medical Image Pre-training |
| RECOMB 2025 | Hierarchical Spatio-Temporal State-Space Modeling for fMRI Analysis | Link | fMRI Analysis | |
| BIBM 2024 | MSVM-UNet: Multi-Scale Vision Mamba UNet for Medical Image Segmentation | Link | Code | 2D Medical Segmentation |
| ICASSP 2025 | OCTAMamba: A State-Space Model Approach for Precision OCTA Vascularization Segmentation | Link | Vessel Segmentation | |
| Information Fusion 2026 | MambaEviScrib: Mamba and Evidence-Guided Consistency Enhance CNN Robustness for Scribble-Supervised Medical Image Segmentation | Link | 2D Medical Segmentation | |
| CMIG 2025 | CT-Mamba: A Hybrid Convolutional State Space Model for Low-Dose CT Denoising | Link | Denoising | |
| International Journal of Imaging Systems and Technology 2025 | Advancing Efficient Brain Tumor Multi‑Class Classification: New Insights From the Vision Mamba Model in Transfer Learning | Link | Medical Image Classification | |
| IEEE RA-L 2025 | MambaXCTrack: Mamba-Based Tracker With SSM Cross-Correlation and Motion Prompt for Ultrasound Needle Tracking | Link | Needle Tracking | |
| WACV 2025 | SAM-Mamba: Mamba Guided SAM Architecture for Generalized Zero-Shot Polyp Segmentation | Link | Code | Polyp Segmentation |
| CVPR 2025 | Unsupervised Foundation Model-Agnostic Slide-Level Representation Learning | Link | Slide-Level Representation Learning | |
| CVPR 2025 | 2DMamba: Efficient State Space Model for Image Representation with Applications on Giga-Pixel Whole Slide Image Classification | Link | Code | Medical Image Classification |
| ICCKE 2024 | Segmentation of Coronary Artery Stenosis in X-ray Angiography using Mamba Model | Link | 2D Medical Segmentation | |
| MICCAI 2025 | Surface Vision Mamba: Leveraging Bidirectional State Space Model for Efficient Spherical Manifold Representation | Link | Code | Spherical Manifold Representation |
| CVPR 2025 | M3amba: Memory Mamba is All You Need for Whole Slide Image Classification | Link | Cancer Subtyping | |
| CVPR 2025 | Cross-Modal Interactive Perception Network with Mamba for Lung Tumor Segmentation in PET-CT Images | Link | Code | Tumor Segmentation |
| ICCV 2025 Oral | GMMamba: Group Masking Mamba for Whole Slide Image Classification | Link | Cancer Subtyping | |
| ICCV 2025 | STDDNet: Harnessing Mamba for Video Polyp Segmentation via Spatial-aligned Temporal Modeling and Discriminative Dynamic Representation Learning | Link | Code | Polyp Segmentation |
| NeurIPS 2025 | Mamba Goes HoME: Hierarchical Soft Mixture-of-Experts for 3D Medical Image Segmentation | Link | Code | 3D Medical Segmentation |
| MICCAI 2025 | HybridMamba: A Dual-domain Mamba for 3D Medical Image Segmentation | Link | 3D Medical Segmentation | |
| MICCAI 2025 | XFMamba: Cross-Fusion Mamba for Multi-View Medical Image Classification | Link | Code | Medical Image Classification |
| MICCAI 2025 | DASMamba: Directional Adaptive Shuffle-Based Visual State-Space Models for Medical Image Restoration | Link | Code | Image Restoration |
| MICCAI 2025 | PolyMamba: Spatial-prior Guided Mamba for Polyp Segmentation | Link | Polyp Segmentation | |
| MICCAI 2025 | CRAViM: Hybrid State-Space Models and Denoising Training for Unpaired Medical Image Synthesis | Link | Code | Image Synthesis |
| MICCAI 2025 | Knowledge-guided Multi-scale Graph Mamba for Whole Slide Image Classification | Link | Cancer Subtyping | |
| MICCAI 2025 | IM-Fuse: A Mamba-based Fusion Block for Brain Tumor Segmentation with Incomplete Modalities | Link | Code | 3D Medical Segmentation |
| MICCAI 2025 | A New Paradigm for Low-dose PET/CT Reconstruction with Mamba-powered Progressive Network | Link | Image Reconstruction | |
| MICCAI 2025 | BrainMT: A Hybrid Mamba-Transformer Architecture for Modeling Long-Range Dependencies in Functional MRI Data | Link | fMRI Analysis | |
| MICCAI 2025 | Dual Correlation-aware Mamba for Microvascular Obstruction Identification in Non-contrast Cine Cardiac Magnetic Resonance | Link | Code | Medical Image Analysis |
| MICCAI 2025 | EndoMamba: An Efficient Foundation Model for Endoscopic Videos via Hierarchical Pre-training | Link | Medical Image Analysis | |
| IEEE TMI 2025 | Diversity-enhanced Collaborative Mamba for Semi-supervised Medical Image Segmentation | Link | 2D & 3D Medical Segmentation | |
| ACM MM 2025 | Unified Medical Image Segmentation with State Space Modeling Snake | Link | 2D Medical Segmentation | |
| IEEE TIP 2025 | COMMA: Coordinate-aware Modulated Mamba Network for 3D Dispersed Vessel Segmentation | Link | Vessel Segmentation | |
| IEEE TMI 2025 | Mamba-Sea: A Mamba-based Framework with Global-to-Local Sequence Augmentation for Generalizable Medical Image Segmentation | Link | Code | 2D Medical Segmentation |
| MICCAI 2025 | MrTrack: Register Mamba for Needle Tracking with Rapid Reciprocating Motion during Ultrasound-Guided Aspiration Biopsy | Link | Needle Tracking | |
| MICCAI 2025 | U-Mamba2: Scaling State Space Models for Dental Anatomy Segmentation in CBCT | Link | 3D Medical Segmentation | |
| BMVC 2025 | CellMamba: Adaptive Mamba for Accurate and Efficient Cell Detection | Link | Cell Detection | |
| MICCAI 2025 Workshop | AMD-Mamba: A Phenotype-Aware Multi-Modal Framework for Robust AMD Prognosis | Link | Risk Prediction | |
| MICCAI 2025 Workshop | PUUMA: Functional MRI Prediction of Gestational Age at Birth and Preterm Risk | Link | fMRI Analysis | |
| ICASSP 2025 | MDN: Mamba-Driven Dualstream Network for Medical Hyperspectral Image Segmentation | Link | 2D Medical Segmentation | |
| BIBM 2025 | MedMamba-YOLO: A Vision State Space Model for Medical Image Detection | Link | Code | Medical Image Analysis |
| BIBM 2025 | ConSSM-GAN: A Contrastive and State-Space Enhanced GAN for MR-to-CT Pelvic Image Translation | Link | Image Synthesis | |
| BIBM 2025 | SWinMamba: Serpentine Window State Space Model for Vascular Segmentation | Link | Vessel Segmentation | |
| BIBM 2025 | DB-MSMUNet: Dual Branch Multi-Scale Mamba UNet for Pancreatic CT Scans Segmentation | Link | 2D Medical Segmentation | |
| BIBM 2025 | Bridging the Perception-Cognition Gap: Re-Engineering SAM2 with Hilbert-Mamba for Robust VLM-Based Medical Diagnosis | Link | Medical Image Analysis | |
| BIBM 2025 | BC-Mamba: Boundary-Aware Contextual CNNs-Mamba for Accurate Ultrasound Image Segmentation | Link | 2D Medical Segmentation | |
| BIBM 2025 | MorphMamba: A Global Context-Aware Mamba for Volumetric Multi-Organ Segmentation | Link | 3D Medical Segmentation | |
| BIBM 2025 | Spatiotemporal Uncertainty-Aware Mamba-Transformer Synergy: Breast Cancer Detection in ABUS | Link | Tumor Segmentation | |
| BIBM 2025 | MM-UNet: Morph Mamba U-Shaped Convolutional Networks for Retinal Vessel Segmentation | Link | Vessel Segmentation | |
| BIBM 2025 | Wavelet Multi-Dimensional and Mamba-Guided Semantic Graph Feature Fusion Network for Glioma Grading | Link | Medical Image Classification | |
| BIBM 2025 | Versatile and Efficient Medical Image Super-Resolution Via Frequency-Gated Mamba | Link | Super-Resolution | |
| BIBM 2025 | E-ViM3: Mamba-3D as Masked Autoencoders for Accurate and Data-Efficient Analysis of Medical Ultrasound Videos | Link | Medical Image Analysis | |
| ICASSP 2025 | SFma-Unet: A Mamba-Based Spatial-Frequency Fusion Network for Medical Image Segmentation | Link | 2D Medical Segmentation | |
| ICASSP 2025 | MTTM: Memory-Augmented with Mamba for 3D Medical Images Analysis | Link | Medical Image Analysis | |
| ICASSP 2025 | Edge-Interaction Mamba Network for MRI Brain Tumor Segmentation | Link | Tumor Segmentation | |
| ICASSP 2025 | PHMamba: Preheating State Space Models with Context-Augmented Features for Medical Image Segmentation | Link | 2D Medical Segmentation | |
| ICASSP 2025 | Causal fMRI-Mamba: Causal State Space Model for Neural Decoding and Brain Task States Recognition | Link | fMRI Analysis | |
| AAAI 2026 | EccoMamba: Enhanced Cross-hierarchical Continuity Orthogonal Mamba for Medical Image Segmentation | Link | 2D Medical Segmentation | |
| AAAI 2026 | HiFi-Mamba: Dual-Stream W-Laplacian Enhanced Mamba for High-Fidelity MRI Reconstruction | Link | Image Reconstruction | |
| AAAI 2026 | Δt-Mamba3D: A Time-Aware Spatio-Temporal State-Space Model for Breast Cancer Risk Prediction | Link | Risk Prediction | |
| AAAI 2026 | Rescind: Countering Image Misconduct in Biomedical Publications with Vision-Language and State-Space Modeling | Link | Medical Image Classification | |
| WACV 2026 | Hymavi: A Hybrid Mamba-Attention Network in Multi-View Framework for Volumetric Medical Image Segmentation | Link | 3D Medical Segmentation | |
| ICASSP 2026 | ConfMamba-SAM: Structured State Space Modeling with Memory-Augmented Prompting for Automatic Brain Lesion Segmentation | Link | 3D Medical Segmentation | |
| ICASSP 2026 | DDMamba: Dual-Scale Constrained Deformable Convolution with Mamba for Medical Image Segmentation | Link | 2D & 3D Medical Segmentation | |
| CVPR 2026 | MambaLiteUNet: Cross-Gated Adaptive Feature Fusion for Robust Skin Lesion Segmentation | Link | 2D Medical Segmentation | |
| CVPR 2026 | VesMamba: 3D Pulmonary Vessel Segmentation from CT images via Mamba with Structural Perception and Scale-aware Filtering | Link | Vessel Segmentation | |
| CVPR 2026 | GeoSemba: Reconstructing State Space Model for Cross Paradigm Representation in Medical Image Segmentation | Link | Code | 2D & 3D Medical Segmentation |
| CVPR 2026 | VEMamba: Efficient Isotropic Reconstruction of Volume Electron Microscopy with Axial-Lateral Consistent Mamba | Link | Code | Image Reconstruction |
| CVPR 2026 Oral | MDCS-MoAME: Multi-directional Composite Scanning with Mixture of Attention and Mamba Experts for Cancer Survival Prediction | Link | Cancer Subtyping/Survival Prediction | |
| ICLR 2026 | BioTamperNet: Affinity-Guided State-Space Model Detecting Tampered Biomedical Images | Link | Medical Image Classification |
Video
| Venue | Paper | Link | Code | Task |
|---|---|---|---|---|
| ECCV 2024 | VideoMamba: State Space Model for Efficient Video Understanding | Link | Code | Video Understanding |
| ICLR 2024 | SSM Meets Video Diffusion Models: Efficient Video Generation with Structured State Spaces | Link | Code | Video Generation |
| IJCV 2026 | Video Mamba Suite: State Space Model as a Versatile Alternative for Video Understanding | Link | Code | Video Understanding |
| CVPR 2024 Workshop | VMRNN: Integrating Vision Mamba and LSTM for Efficient and Accurate Spatiotemporal Forecasting | Link | Code | Spatiotemporal Forecasting |
| ICCV 2025 | Snakes and Ladders: Two Steps Up for VideoMamba | Link | Code | Video Understanding |
| AAAI 2025 | RhythmMamba: Fast Remote Physiological Measurement with Arbitrary Length Videos | Link | Code | Remote Photoplethysmography |
| NeurIPS 2024 | VFIMamba: Video Frame Interpolation with State Space Models | Link | Code | Frame Interpolation |
| ECCV 2024 | VideoMamba: Spatio-Temporal Selective State Space Model | Link | Code | Action Recognition |
| ACM MM 2024 Oral | RainMamba: Enhanced Locality Learning with State Space Models for Video Deraining | Link | Code | Deraining |
| ACM MM 2024 Oral | MambaTrack: A Simple Baseline for Multiple Object Tracking with State Space Model | Link | Multi-Object Tracking | |
| Computers and Electronics in Agriculture 2025 | FMRFT: Fusion Mamba and DETR for Query Time Sequence Intersection Fish Tracking | Link | Multi-Object Tracking | |
| IEEE JSTAR 2024 | TrackingMamba: Visual State Space Model for Object Tracking | Link | Code | Object Tracking |
| CCBR 2024 | PhysMamba: Efficient Remote Physiological Measurement with SlowFast Temporal Difference Mamba | Link | Code | Remote Photoplethysmography |
| NeurIPS 2024 | MambaSCI: Efficient Mamba-UNet for Quad-Bayer Patterned Video Snapshot Compressive Imaging | Link | Code | Snapshot Compressive Imaging |
| NeurIPS 2024 | Toward Dynamic Non-Line-of-Sight Imaging with Mamba Enforced Temporal Consistency | Link | Dynamic Reconstruction | |
| ACM MM 2024 | Object-Level Pseudo-3D Lifting for Distance-Aware Tracking | Link | Multi-Object Tracking | |
| AAAI 2025 | Manta: Enhancing Mamba for Few-Shot Action Recognition of Long Sub-Sequence | Link | Code | Action Recognition |
| AAAI 2025 | Exploring Enhanced Contextual Information for Video-Level Object Tracking | Link | Code | Object Tracking |
| AAAI 2025 | Robust Tracking via Mamba-based Context-aware Token Learning | Link | Code | Object Tracking |
| IROS 2025 | MambaNUT: Nighttime UAV Tracking via Mamba-based Adaptive Curriculum Learning | Link | Object Tracking | |
| AAAI 2025 | Efficient Self-Supervised Video Hashing with Selective State Spaces | Link | Code | Hashing |
| IEEE TMM 2026 | STNMamba: Mamba-based Spatial-Temporal Normality Learning for Video Anomaly Detection | Link | Anomaly Detection | |
| CVPR 2025 | MambaVO: Deep Visual Odometry Based on Sequential Matching Refinement and Training Smoothing | Link | Visual Odometry | |
| NeurIPS 2024 | Slot State Space Models | Link | Code | Video Understanding |
| IEEE TPAMI 2025 | MADiff: Motion-Aware Mamba Diffusion Models for Hand Trajectory Prediction on Egocentric Videos | Link | Trajectory Prediction | |
| CVPR 2025 | MANTA: Diffusion Mamba for Efficient and Effective Stochastic Long-term Dense Anticipation | Link | Action Anticipation | |
| CVPR 2025 | Event-based Video Super-Resolution via State Space Models | Link | Super-Resolution | |
| CVPR 2025 | LC-Mamba: Local and Continuous Mamba with Shifted Windows for Frame Interpolation | Link | Frame Interpolation | |
| ICCV 2025 | Vamba: Understanding Hour-Long Videos with Hybrid Mamba-Transformers | Link | Video Understanding | |
| ICCV 2025 | VSRM: A Robust Mamba-Based Framework for Video Super-Resolution | Link | Super-Resolution | |
| ICCV 2025 | EVDM: Event-based Real-World Video Deblurring with Mamba | Link | Deblurring | |
| ICCV 2025 | EgoMusic-Driven Human Dance Motion Estimation with Skeleton Mamba | Link | Human Motion Estimation | |
| ICCV 2025 | PS-Mamba: Spatial-Temporal Graph Mamba for Pose Sequence Refinement | Link | 3D Human Pose Estimation | |
| ICML 2025 | MoMa: Modulating Mamba for Adapting Image Foundation Models to Video Recognition | Link | Action Recognition | |
| ACM MM 2025 | UMSD: High Realism Motion Style Transfer via Unified Mamba-based Diffusion | Link | Motion Style Transfer | |
| CVPR 2025 Highlight | Learning Phase Distortion with Selective State Space Models for Video Turbulence Mitigation | Link | Video Restoration | |
| CVPR 2025 Oral | Semi-Supervised State-Space Model with Dynamic Stacking Filter for Real-World Video Deraining | Link | Video Restoration | |
| CVPR 2025 | Self-supervised ControlNet with Spatio-Temporal Mamba for Real-world Video Super-resolution | Link | Super-Resolution | |
| ICCV 2025 | PRE-Mamba: A 4D State Space Model for Ultra-High-Frequent Event Camera Deraining | Link | Video Restoration | |
| ICCV 2025 | High-Resolution Spatiotemporal Modeling with Global-Local State Space Models for Video-Based Human Pose Estimation | Link | Human Pose Estimation | |
| NeurIPS 2025 | MVSMamba: Multi-View Stereo with State Space Model | Link | Multi-View Stereo | |
| ICASSP 2025 | MambaMOT: State-Space Model as Motion Predictor for Multi-Object Tracking | Link | Multi-Object Tracking | |
| ICASSP 2025 | MambaTrack: Exploiting Dual-Enhancement for Night UAV Tracking | Link | Object Tracking | |
| CVPR 2025 Workshop | SportMamba: Adaptive Non-Linear Multi-Object Tracking with State Space Models for Team Sports | Link | Multi-Object Tracking | |
| CVPR 2025 Workshop | Dyadic Mamba: Long-term Dyadic Human Motion Synthesis | Link | Motion Generation | |
| ICASSP 2025 | Sign-Mamba: Advanced Mamba-Based Sign Language Generation | Link | Motion Generation | |
| ICASSP 2025 | DSSM: Dual State Space Model for Human Motions Generation | Link | Motion Generation | |
| AAAI 2026 | MambaOVSR: Multiscale Fusion with Global Motion Modeling for Chinese Opera Video Super-Resolution | Link | Super-Resolution | |
| AAAI 2026 | State-Space Hierarchical Compression with Gated Attention and Learnable Sampling for Hour-Long Video Understanding in Large Multimodal Models | Link | Code | Video Understanding |
| AAAI 2026 | Backtrace Mamba: Reviving Critical Temporal Contexts via Hierarchical Memory Compression for Online Action Detection | Link | Action Detection | |
| AAAI 2026 | DeformTrace: A Deformable State Space Model with Relay Tokens for Temporal Forgery Localization | Link | Temporal Forgery Localization | |
| ICLR 2026 | ConvT3: Structured State Kernels for Convolutional State Space Models | Link | Video Generation | |
| ICLR 2026 | Trajectory-aware Shifted State Space Models for Online Video Super-Resolution | Link | Code | Super-Resolution |
| CVPR 2026 | RS-SSM: Refining Forgotten Specifics in State Space Model for Video Semantic Segmentation | Link | Video Semantic Segmentation | |
| CVPR 2026 | M4V: Multimodal Mamba for Efficient Text-to-Video Generation | Link | Video Generation | |
| CVPR 2026 | HieraMamba: Video Temporal Grounding via Hierarchical Anchor-Mamba Pooling | Link | Temporal Grounding | |
| CVPR 2026 | MS-Temba: Multi-Scale Temporal Mamba for Understanding Long Untrimmed Videos | Link | Action Recognition | |
| CVPR 2026 | EgoFlow: Gradient-Guided Flow Matching for Egocentric 6DoF Object Motion Generation | Link | Motion Generation | |
| CVPR 2026 | Gamba: Mamba-based Graph Convolutional Network with Dynamic Graph Topology Learning for Action Recognition | Link | Action Recognition | |
| CVPR 2026 | When Transformers Meet Mamba: A Hybrid Transformer-Mamba Network for Video Object Detection | Link | Video Object Detection | |
| CVPR 2026 Findings | MVSSM: Motion-aware Visual State Space Model for Efficient Video Deblurring | Link | Code | Video Deblurring |
| ICML 2026 | StructMamPose: From Sequential Perception to Structural Reasoning for 3D Human Pose Estimation | Link | 3D Human Pose Estimation | |
| ECCV 2026 | MASS: Motion-Aligned Selective Scan for Refinement in Flow-Based Video Frame Interpolation | Link | Video Frame Interpolation |
Point Cloud
| Venue | Paper | Link | Code | Task |
|---|---|---|---|---|
| NeurIPS 2024 | PointMamba: A Simple State Space Model for Point Cloud Analysis | Link | Code | Classification, Part Segmentation |
| CVPR 2024 | State Space Models for Event Cameras | Link | Code | Object Detection |
| ACM MM 2024 | MambaMOS: LiDAR-based 3D Moving Object Segmentation with Motion-aware State Space Model | Link | Code | Object Segmentation |
| ACM MM 2024 | Mamba3D: Enhancing Local Features for 3D Point Cloud Analysis via State Space Model | Link | Code | Classification, Part Segmentation |
| NeurIPS 2024 | LCM: Locally Constrained Compact Point Cloud Model for Masked Point Modeling | Link | Code | Classification, Part Segmentation, 3D Object Detection |
| NeurIPS 2024 | Voxel Mamba: Group-Free State Space Models for Point Cloud based 3D Object Detection | Link | Code | 3D Object Detection |
| NeurIPS 2024 | 3DET-Mamba: Causal Sequence Modelling for End-to-End 3D Object Detection | Link | 3D Object Detection | |
| ICIP 2024 | Mamba-PCGC: Mamba-Based Point Cloud Geometry Compression | Link | Geometry Compression | |
| NeurIPS 2024 | LION: Linear Group RNN for 3D Object Detection in Point Clouds | Link | Code | 3D Object Detection |
| AAAI 2025 | SMamba: Sparse Mamba for Event-based Object Detection | Link | Code | Object Detection |
| AAAI 2025 | Point Cloud Mamba: Point Cloud Learning via State Space Model | Link | Code | Classification, Segmentation |
| AAAI 2025 | 3DMambaIPF: A State Space Model for Iterative Point Cloud Filtering | Link | Point Cloud Filtering | |
| IEEE TPAMI 2025 | Rethinking Efficient and Effective Point-based Networks for Event Camera Classification and Regression | Link | Event Camera Classification | |
| CVPR 2025 | MAMBA4D: Efficient Long-Sequence Point Cloud Video Understanding with Disentangled Spatial-Temporal State Space Models | Link | Point Cloud Video Understanding | |
| AAAI 2025 | Pamba: Enhancing Global Interaction in Point Clouds via State Space Model | Link | Point Cloud Analysis | |
| ROBIO 2024 | MV-MOS: Multi-View Feature Fusion for 3D Moving Object Segmentation | Link | Code | Object Segmentation |
| IEEE TCSVT 2025 | MambaEVT: Event Stream-Based Visual Object Tracking Using State Space Model | Link | Code | Object Tracking |
| IEEE RA-L 2025 | OMEGA: Efficient Occlusion-Aware Navigation for Air-Ground Robots in Dynamic Environments via State Space Model | Link | Robot Navigation | |
| ACM MM Asia 2024 | SpikMamba: When SNN meets Mamba in Event-based Human Action Recognition | Link | Code | Action Recognition |
| Communications in Transportation Research 2025 | MetaSSC: Enhancing 3D Semantic Scene Completion for Autonomous Driving through Meta-Learning and Long-sequence Modeling | Link | 3D Semantic Scene Completion | |
| CVPR 2025 | WeatherGen: A Unified Diverse Weather Generator for LiDAR Point Clouds via Spider Mamba Diffusion | Link | Code | Point Cloud Generation |
| CVPR 2025 | Spectral Informed Mamba for Robust Point Cloud Processing | Link | Point Cloud Analysis | |
| ICCV 2025 | Efficient Spiking Point Mamba for Point Cloud Analysis | Link | Point Cloud Analysis | |
| ICCV 2025 | StruMamba3D: Exploring Structural Mamba for Self-supervised Point Cloud Representation Learning | Link | Self-supervised Learning | |
| ICLR 2025 | MamBEV: Enabling State Space Models to Learn Birds-Eye-View Representations | Link | 3D Object Detection | |
| ICLR 2025 | State Space Model Meets Transformer: A New Paradigm for 3D Object Detection | Link | 3D Object Detection | |
| CVPR 2025 | UniMamba: Unified Spatial-Channel Representation Learning with Group-Efficient Mamba for LiDAR-based 3D Object Detection | Link | 3D Object Detection | |
| CVPR 2025 | PMA: Towards Parameter-Efficient Point Cloud Understanding via Point Mamba Adapter | Link | Point Cloud Analysis | |
| ICCV 2025 | UST-SSM: Unified Spatio-Temporal State Space Models for Point Cloud Video Modeling | Link | Point Cloud Analysis | |
| AAAI 2026 | Seeing in Double: Dual-Granularity BEV Segmentation via Mamba-Driven Alignment and Polar-Decoupled Experts | Link | BEV Segmentation | |
| AAAI 2026 | DAPointMamba: Domain Adaptive Point Mamba for Point Cloud Completion | Link | Point Cloud Completion | |
| AAAI 2026 | CloudMamba: Grouped Selective State Spaces for Point Cloud Analysis | Link | Point Cloud Analysis | |
| AAAI 2026 | BeyondSparse: Facilitating Mamba to Enhance Cross-Domain 3D Semantic Segmentation in Adverse Weather | Link | Semantic Segmentation | |
| AAAI 2026 | WinMamba: Multi-Scale Shifted Windows in State Space Model for 3D Object Detection | Link | 3D Object Detection | |
| WACV 2026 | Towards Streaming LiDAR Object Detection with Point Clouds as Egocentric Sequences | Link | 3D Object Detection | |
| WACV 2026 | MEGA-PCC: A Mamba-based Efficient Approach for Joint Geometry and Attribute Point Cloud Compression | Link | Point Cloud Compression | |
| WACV 2026 | SSMRadNet: A Sample-wise State-Space Framework for Efficient and Ultra-Light Radar Segmentation and Object Detection | Link | Object Detection | |
| WACV 2026 | milliMamba: Specular-Aware Human Pose Estimation via Dual mmWave Radar with Multi-Frame Mamba Fusion | Link | 3D Human Pose Estimation | |
| CVPR 2026 | GEM: Generating LiDAR World Model via Deformable Mamba | Link | LiDAR World Model | |
| CVPR 2026 | MARSS: Radar Semantic Segmentation via Modular Attention and State Space Models | Link | Semantic Segmentation | |
| CVPR 2026 | Mamba Learns in Context: Structure-Aware Domain Generalization for Multi-Task Point Cloud Understanding | Link | Point Cloud Understanding | |
| ICML 2026 | NeuroMamba: A Universal Spatiotemporal Module for Robust Perception in Degraded Sensory Streams | Link | Point Cloud Analysis | |
| ICLR 2026 | Fore-Mamba3D: Mamba-based Foreground-Enhanced Encoding for 3D Object Detection | Link | Code | 3D Object Detection |
| ICLR 2026 | 3DSMT: A Hybrid Spiking Mamba-Transformer for Point Cloud Analysis | Link | Point Cloud Analysis | |
| ICLR 2026 | DriveMamba: Task-Centric Scalable State Space Model for Efficient End-to-End Autonomous Driving | Link | Autonomous Driving | |
| ICLR 2026 | Point-Focused Attention Meets Context-Scan State Space: Robust Biological Visual Perception for Point Cloud Representation | Link | Point Cloud Analysis | |
| ICASSP 2026 | STNID: A Spatiotemporal Mamba-Based Neural Implicit Dynamics Model for Point Cloud Forecasting | Link | Point Cloud Forecasting |
Multi-Modal
| Venue | Paper | Link | Code | Task | Modality |
|---|---|---|---|---|---|
| Information Fusion 2025 | Pan-Mamba: Effective pan-sharpening with State Space Model | Link | Code | Pansharpening | HISR Images & LRMS Images |
| ECCV 2024 | InstructGIE: Towards Generalizable Image Editing | Link | Code | Image Editing | Image & Text |
| ECCV 2024 | Motion Mamba: Efficient and Long Sequence Motion Generation with Hierarchical and Bidirectional Selective SSM | Link | Code | Text-to-Motion Generation | Motion & Text |
| NeurIPS 2024 Workshop | VL-Mamba: Exploring State Space Models for Multimodal Learning | Link | Code | MLLM Tasks | Image & Text |
| Neural Computing and Applications 2026 | Music to Dance as Language Translation using Sequence Models | Link | Code | Dance Generation | Motion & Audio |
| NeurIPS 2024 | MambaTalk: Efficient Holistic Gesture Synthesis with Selective State Space Models | Link | Gesture Generation | Motion & Audio | |
| ECCV 2024 | ReMamber: Referring Image Segmentation with Mamba Twister | Link | Code | Referring Image Segmentation | Image & Text |
| BIBM 2025 | SurvMamba: State Space Model with Multi-grained Multi-modal Interaction for Survival Prediction | Link | Cancer Subtyping/Survival Prediction | WSIs & Gene | |
| WACV 2025 | Sigma: Siamese Mamba Network for Multi-Modal Semantic Segmentation | Link | Code | Semantic Segmentation | RGB Images & Depth Images / Thermal Images |
| IEEE TGRS 2024 | Efficient Remote Sensing Image Fusion With State Space Model | Link | Code | Pansharpening | HISR Images & LRMS Images |
| PRCV 2024 | Mamba-FETrack: Frame-Event Tracking via State Space Model | Link | Code | RGB-Event Tracking | RGB Frames & Event Data |
| IEEE GRSL 2024 | RSCaMa: Remote Sensing Image Change Captioning with State Space Model | Link | Code | Image Captioning | Remote Sensing Images & Text |
| MIA 2025 | MMR-Mamba: Multi-modal MRI reconstruction with Mamba and spatial-frequency information fusion | Link | Image Fusion | Multi-Contrast MRI | |
| IEEE TIP 2025 | S4Fusion: Saliency-Aware Selective State Space Model for Infrared and Visible Image Fusion | Link | Image Fusion | RGB Images & Infrared Images | |
| NeurIPS 2024 | Meteor: Mamba-based Traversal of Rationale for Large Language and Vision Models | Link | Code | MLLM Tasks | Image & Text |
| NeurIPS 2024 | Coupled Mamba: Enhanced Multi-modal Fusion with Coupled State Space Model | Link | Multi-modal Sentiment Analysis | Video & Text & Audio | |
| NeurIPS 2024 | RoboMamba: Multimodal State Space Model for Efficient Robot Reasoning and Manipulation | Link | Code | Robot Reasoning and Manipulation | Image & Text |
| ACM MM 2024 | MambaGesture: Enhancing Co-Speech Gesture Generation with Mamba and Disentangled Multi-Modality Fusion | Link | Gesture Generation | Motion & Audio | |
| ITSC 2024 | MambaST: A Plug-and-Play Cross-Spectral Spatial-Temporal Fuser for Efficient Pedestrian Detection | Link | Code | Pedestrian Detection | RGB Images & Thermal Images |
| ACML 2024 | ColorMamba: Towards High-quality NIR-to-RGB Spectral Translation with Mamba | Link | Code | NIR-to-RGB Translation | RGB Images & NIR Images |
| ACM TOMCCAP 2025 | JambaTalk: Speech-driven 3D Talking Head Generation based on a Hybrid Transformer-Mamba Model | Link | 3D Talking Head Generation | Motion & Audio | |
| IEEE TGRS 2024 | Mask-Guided Mamba Fusion for Drone-based Visible-Infrared Vehicle Detection | Link | Object Detection | RGB Images & Infrared Images | |
| IEEE TGRS 2024 | Joint Classification of Hyperspectral and LiDAR Data Based on Mamba | Link | Code | Classification | HSI & LiDAR Points |
| MICCAI 2024 | LM-UNet: Whole-Body PET-CT Lesion Segmentation with Dual-Modality-Based Annotations Driven by Latent Mamba U-Net | Link | Code | 3D Medical Segmentation | PET Images & CT Images |
| ICLR 2025 | EMMA: Empowering Multi-modal Mamba with Structural and Hierarchical Alignment | Link | MLLM Tasks | Image & Text | |
| ISBI 2025 | R2Gen-Mamba: A Selective State Space Model for Radiology Report Generation | Link | Code | Radiology Report Generation | Image & Text |
| BDAI 2025 | MSCrackMamba: Leveraging Vision Mamba for Crack Detection in Fused Multispectral Imagery | Link | Crack Detection | RGB Images & Infrared Images | |
| CVIP 2025 | DiM-Gestor: Co-Speech Gesture Generation with Adaptive Layer Normalization Mamba-2 | Link | Code | Gesture Generation | Motion & Audio |
| IEEE GRSL 2024 | A Mamba-Diffusion Framework for Multimodal Remote Sensing Image Semantic Segmentation | Link | Code | Semantic Segmentation | RGB Images & SAR Images |
| IEEE TIV 2024 | SeqMamba-MPR: A Spatial-Temporal Mamba Network for Place Recognition Using Sequential Multi-Modal Data | Link | Place Recognition | RGB Images & LiDAR Points | |
| IEEE GRSL 2024 | S2CrossMamba: Spatial–Spectral Cross-Mamba for Multimodal Remote Sensing Image Classification | Link | Code | Classification | HISR Images & LRMS Images |
| Scientific Reports 2024 | ReMamba: a hybrid CNN-Mamba aggregation network for visible-infrared person re-identification | Link | Visible-infrared Re-identification | RGB Images & Infrared Images | |
| IEEE TCSVT 2024 | MDNet: Mamba-Effective Diffusion-Distillation Network for RGB-Thermal Urban Dense Prediction | Link | Code | Dense Prediction | RGB Images & Thermal Images |
| CVPR 2025 | AlignMamba: Enhancing Multimodal Mamba with Local and Global Cross-modal Alignment | Link | Multi-Modality Fusion | Video & Text & Audio | |
| AAAI 2025 | LOMA: Language-assisted Semantic Occupancy Network via Triplane Mamba | Link | Occupancy Prediction | Image & Text | |
| AAAI 2025 | MambaPro: Multi-Modal Object Re-Identification with Mamba Aggregation and Synergistic Prompt | Link | Code | Object Re-Identification | RGB Images & Near Infrared Images & Thermal Infrared Images |
| AAAI 2025 | Light-T2M: A Lightweight and Fast Model for Text-to-motion Generation | Link | Code | Text-to-Motion Generation | Motion & Text |
| AAAI 2025 | Exploiting Multimodal Spatial-temporal Patterns for Video Object Tracking | Link | Code | Object Tracking | RGB Videos & TIR/Depth/Event |
| ICASSP 2025 | Trusted Mamba Contrastive Network for Multi-View Clustering | Link | Multi-View Clustering | Image & Text | |
| AAAI 2025 | H-MBA: Hierarchical MamBa Adaptation for Multi-Modal Video Understanding in Autonomous Driving | Link | Video Understanding | Video & Text | |
| AAAI 2025 | Skip Mamba Diffusion for Monocular 3D Semantic Scene Completion | Link | Code | 3D Semantic Scene Completion | RGB Images & LiDAR Points |
| IEEE TMM 2025 | AVS-Mamba: Exploring Temporal and Multi-modal Mamba for Audio-Visual Segmentation | Link | Code | Audio-Visual Segmentation | Video & Audio |
| ICCV 2025 | LiT: Delving into a Simplified Linear Diffusion Transformer for Image Generation | Link | Code | Image Generation | Image & Text |
| Information Fusion 2025 | An efficient cross-view image fusion method based on selected state space and hashing for promoting urban perception | Link | Geo-Localization | Street-view Images & Aerial-view Images | |
| AAAI 2025 | Cobra: Extending Mamba to Multi-Modal Large Language Model for Efficient Inference | Link | Code | MLLM Tasks | Image & Text |
| IEEE TMM 2025 | Fusion-Mamba for Cross-modality Object Detection | Link | Object Detection | RGB Images & Infrared Images | |
| Visual Intelligence 2025 | FusionMamba: Dynamic Feature Enhancement for Multimodal Image Fusion with Mamba | Link | Code | Image Fusion | Multi-Modal Images |
| IEEE TIP 2025 | Text-controlled Motion Mamba: Text-Instructed Temporal Grounding of Human Motion | Link | Temporal Grounding | Motion & Text | |
| IEEE TCSVT 2025 | CFMW: Cross-modality Fusion Mamba for Robust Object Detection under Adverse Weather with Multimodal Camera | Link | Code | Object Detection | RGB Images & Infrared Images |
| AAAI 2025 | RGBT Tracking via All-layer Multimodal Interactions with Progressive Fusion Mamba | Link | RGB-T Tracking | RGB Images & Thermal Images | |
| IEEE TCSVT 2025 | MambaVT: Spatio-Temporal Contextual Modeling for robust RGB-T Tracking | Link | RGB-T Tracking | RGB Images & Thermal Images | |
| IEEE JBHI 2026 | R2GenCSR: Mining Contextual and Residual Information for LLMs-based Radiology Report Generation | Link | Radiology Report Generation | Image & Text | |
| CVPR 2025 | OccMamba: Semantic Occupancy Prediction with State Space Models | Link | Semantic Occupancy Prediction | RGB Images & LiDAR Points | |
| AAAI 2025 | MUSE: Mamba is Efficient Multi-scale Learner for Text-video Retrieval | Link | Text-Video Retrieval | Video & Text | |
| IEEE TETCI 2026 | DualKanbaFormer: An Efficient Selective Sparse Framework for Multimodal Aspect-Based Sentiment Analysis | Link | Multi-modal Sentiment Analysis | Image & Text | |
| IROS 2025 | MambaPlace:Text-to-Point-Cloud Cross-Modal Place Recognition with Attention Mamba Mechanisms | Link | Code | Place Recognition | 3D Point Cloud & Text |
| IEEE TCSVT 2025 | Shuffle Mamba: State Space Models with Random Shuffle for Multi-Modal Image Fusion | Link | Image Fusion | RGB Images & Infrared Images | |
| ICASSP 2025 | Mamba Fusion: Learning Actions Through Questioning | Link | Code | Action Anticipation | Video & Text |
| ICASSP 2025 | Mamba-YOLO-World: Marrying YOLO-World with Mamba for Open-Vocabulary Detection | Link | Code | Open-Vocabulary Detection | Image & Text |
| ADMA 2025 | Mamba-Enhanced Text-Audio-Video Alignment Network for Emotion Recognition in Conversations | Link | Code | Emotion Recognition | Video & Text & Audio |
| IROS 2025 | GraspMamba: A Mamba-based Language-driven Grasp Detection Framework with Hierarchical Feature Learning | Link | Code | Grasp Detection | Image & Text |
| ACM MM Asia 2024 | LMHaze: Intensity-aware Image Dehazing with a Large-scale Multi-intensity Real Haze Dataset | Link | Dehazing | Image & Text | |
| Neurocomputing 2025 | MambaSOD: Dual Mamba-Driven Cross-Modal Fusion Network for RGB-D Salient Object Detection | Link | Code | Salient Object Detection | RGB Images & Depth Images |
| ACM MM Workshop 2024 | Moyun: A Diffusion-Based Model for Style-Specific Chinese Calligraphy Generation | Link | Style-Specific Chinese Calligraphy Generation | Image & Text | |
| ICASSP 2025 | DepMamba: Progressive Fusion Mamba for Multimodal Depression Detection | Link | Code | Multi-modal Depression Detection | Video & Audio |
| CVPR 2025 | CXPMRG-Bench: Pre-training and Benchmarking for X-ray Medical Report Generation on CheXpert Plus Dataset | Link | Code | Medical Report Generation | Image & Text |
| EMNLP 2024 | Shaking Up VLMs: Comparing Transformers and Structured State Space Models for Vision & Language Modeling | Link | Code | MLLM Tasks | Image & Text |
| EMNLP 2025 Findings | LongLLaVA: Scaling Multi-modal LLMs to 1000 Images Efficiently via a Hybrid Mamba-Transformer Model | Link | Code | MLLM Tasks | Image & Text |
| IROS 2025 | Mamba Policy: Towards Efficient 3D Diffusion Policy with Hybrid Selective State Models | Link | Code | Robot Manipulation | 3D Point Cloud & Text |
| CVPR 2025 | MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking | Link | Vision-Language Tracking | RGB Images & Text | |
| CVPR 2025 | LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity | Link | Text-to-Video Generation | Video & Text | |
| IEEE TPAMI 2025 | OccScene: Semantic Occupancy-based Cross-task Mutual Learning for 3D Scene Generation | Link | 3D Scene Generation | RGB Images & Occupancy Grids | |
| CVPR 2025 | Completion as Enhancement: A Degradation-Aware Selective Image Guided Mamba Fusion Network for Depth Completion | Link | Depth Completion | RGB Images & Depth Images | |
| CVPR 2025 | Exploring Historical Information for RGBE Visual Tracking with Mamba | Link | RGB-Event Tracking | RGB Frames & Event Data | |
| ICCV 2025 | Mamba-3VL: Taming State Space Model for 3D Vision Language Learning | Link | 3D Vision-Language Learning | 3D Point Cloud & Text | |
| IJCAI 2025 | RRG-Mamba: Efficient Radiology Report Generation with State Space Model | Link | Radiology Report Generation | Image & Text | |
| CVPR 2025 | BIMBA: Selective-Scan Compression for Long-Range Video Question Answering | Link | Video QA | Video & Text | |
| ICCV 2025 | MUG: Pseudo Labeling Augmented Audio-Visual Mamba Network for Audio-Visual Video Parsing | Link | Audio-Visual Parsing | Video & Audio | |
| ICCV 2025 | End-to-End Multi-Modal Diffusion Mamba | Link | Multi-modal Generation | Image & Text | |
| ACM MM 2025 | LEAF-Mamba: Local Emphatic and Adaptive Fusion State Space Model for RGB-D Salient Object Detection | Link | Salient Object Detection | RGB Images & Depth Images | |
| ACM MM 2025 | LIDAR: Lightweight Adaptive Cue-Aware Fusion Vision Mamba for Multimodal Segmentation of Structural Cracks | Link | Crack Segmentation | Multi-Modal | |
| NeurIPS 2025 | SaFiRe: Saccade-Fixation Reiteration with Mamba for Referring Image Segmentation | Link | Referring Image Segmentation | Image & Text | |
| ICCV 2025 | MambaFusion: Height-Fidelity Dense Global Fusion for Multi-modal 3D Object Detection | Link | 3D Object Detection | RGB Images & LiDAR Points | |
| ICCV 2025 | WaveMamba: Wavelet-Driven Mamba Fusion for RGB-Infrared Object Detection | Link | Object Detection | RGB Images & Infrared Images | |
| AAAI 2026 | MambaSeg: Harnessing Mamba for Accurate and Efficient Image-Event Semantic Segmentation | Link | Semantic Segmentation | RGB Frames & Event Data | |
| AAAI 2026 | Self-supervised Multiplex Consensus Mamba for General Image Fusion | Link | Image Fusion | Multi-Modal Images | |
| AAAI 2026 | Exploiting All Mamba Fusion for Efficient RGB-D Tracking | Link | Object Tracking | RGB Images & Depth Images | |
| AAAI 2026 | KineST: A Kinematics-guided Spatiotemporal State Space Model for Human Motion Tracking from Sparse Signals | Link | Motion Tracking | Motion & Sparse Signals | |
| WACV 2026 | Not Like Transformers: Drop the Beat Representation for Dance Generation with Mamba-Based Diffusion Model | Link | Dance Generation | Motion & Audio | |
| CVPR 2026 | TimeViper: A Hybrid Mamba-Transformer Vision-Language Model for Efficient Long Video Understanding | Link | Video Understanding | Video & Text | |
| CVPR 2026 | Echoes Over Time: Unlocking Length Generalization in Video-to-Audio Generation Models | Link | Video-to-Audio Generation | Video & Audio | |
| CVPR 2026 | RI-Mamba: Rotation-Invariant Mamba for Robust Text-to-Shape Retrieval | Link | Code | Text-to-Shape Retrieval | 3D Point Cloud & Text |
| CVPR 2026 | VIMCAN: Visual-Inertial 3D Human Pose Estimation with Hybrid Mamba-Cross-Attention Network | Link | Code | 3D Human Pose Estimation | Image & IMU |
| CVPR 2026 | AIMDepth: Asymmetric Image-Event Mamba for Monocular Depth Estimation | Link | Depth Estimation | RGB Frames & Event Data | |
| ICASSP 2026 | MAPD-Mamba: Modality-Adaptive Perception-Driven Mamba Fusion Network | Link | Multimodal Fusion | Multi-Modal |
Others
| Venue | Paper | Link | Code | Task |
|---|---|---|---|---|
| PLOS ONE 2025 | Res-VMamba: Fine-Grained Food Category Visual Classification Using Selective State Space Models with Deep Residual Learning | Link | Code | Food Classification |
| ICRA 2025 | Motion-Guided Dual-Camera Tracker for Low-Cost Skill Evaluation of Gastric Endoscopy | Link | Code | Endoscope Tip Tracking |
| ICLR 2025 | Sports-Traj: A Unified Trajectory Generation Model for Multi-Agent Movement in Sports | Link | Code | Trajectory Generation |
| Sleep 2025 | Mamba-based deep learning approach for sleep staging on a wireless multimodal wearable system without electroencephalography | Link | Sleep Staging | |
| Scientific Reports 2025 | Optimising TinyML with Quantization and Distillation of Transformer and Mamba Models for Indoor Localisation on Edge Devices | Link | Code | Indoor Localization |
| Journal of Physics: Photonics 2025 | Bidirectional Mamba state-space model for anomalous diffusion | Link | Code | Anomalous Diffusion Analysis |
| ICCC 2024 | ST-Mamba: Spatial-Temporal Mamba for Traffic Flow Estimation Recovery with Missing Data | Link | Traffic Flow Estimation | |
| ICCV 2025 | MamTiff-CAD: Multi-Scale Latent Diffusion with Mamba+ for Complex Parametric Sequence Generation | Link | CAD Sequence Generation | |
| CVPR 2025 | Trajectory Mamba: Efficient Attention-Mamba Forecasting Model Based on Selective SSM | Link | Trajectory Prediction | |
| CVPR 2025 Workshop | U-Shape Mamba: State Space Model for faster diffusion | Link | Code | Diffusion Acceleration |
| ICASSP 2025 | Stochastic-Aware Mamba Diffusion for Pedestrian Trajectory Prediction | Link | Trajectory Prediction | |
| AAAI 2026 | Mamba-Driven Multi-View Discriminative Clustering via Global-Local Cross-View Sequence Modeling | Link | Multi-View Clustering | |
| CVPR 2026 | FoSS: Modeling Long-Range Dependencies and Multimodal Uncertainty in Trajectory Prediction via Fourier-State Space Integration | Link | Trajectory Prediction |
Valuable Insights
| Venue | Paper | Link |
|---|---|---|
| ACL 2025 | The Hidden Attention of Mamba Models | Link |
| CVPR 2025 | MambaOut: Do We Really Need Mamba for Vision? | Link |
| NeurIPS 2024 | Demystify Mamba in Vision: A Linear Attention Perspective | Link |
| ICLR 2025 | A Unified Implicit Attention Formulation for Gated-Linear Recurrent Sequence Models | Link |
| NeurIPS 2024 | MambaLRP: Explaining Selective State Space Sequence Models | Link |
| NeurIPS 2025 | TRUST: Test-Time Refinement using Uncertainty-Guided SSM Traverses | Link |
| CVPR 2025 Workshop | Towards Evaluating the Robustness of Visual State Space Models | Link |
| ICML 2025 Workshop | State Space Models: A Naturally Robust Alternative to Transformers in Computer Vision | Link |
| ICLR 2025 Spotlight | Demystifying the Token Dynamics of Deep Selective State Space Models | Link |
| ACL 2025 | Mamba Knockout for Unraveling Factual Information Flow | Link |
| CVPR 2026 Findings | Exemplar-Free Continual Learning for State Space Models | Link |
| CVPR 2026 | RNN as Linear Transformer: A Closer Investigation into Representational Potentials of Visual Mamba Models | Link |
Other Domains
Reinforcement Learning
| Venue | Paper | Link | Code |
|---|---|---|---|
| IROS 2024 | Proprioception Is All You Need: Terrain Classification for Boreal Forests | Link | Code |
| IEEE Access 2025 | Mamba as a motion encoder for robotic imitation learning | Link | |
| ICCMA 2024 | Context Aware Mamba-based Reinforcement Learning for Social Robot Navigation | Link | |
| NeurIPS 2024 | Is Mamba Compatible with Trajectory Optimization in Offline Reinforcement Learning? | Link | |
| NeurIPS 2024 | Decision Mamba: A Multi-Grained State Space Model with Self-Evolution Regularization for Offline RL | Link | |
| CoRL 2024 | MaIL: Improving Imitation Learning with Selective State Space Models | Link | Code |
| IEEE TMRB 2024 | Visuomotor Policy Learning for Task Automation of Surgical Robot | Link | Code |
| NeurIPS 2024 | Decision Mamba: Reinforcement Learning via Hybrid Selective Sequence Modeling | Link | |
| ICRA 2026 | DiSPo: Diffusion-SSM based Policy Learning for Coarse-to-Fine Action Abstraction | Link | |
| ICLR 2025 | Drama: Mamba-Enabled Model-Based Reinforcement Learning Is Sample and Parameter Efficient | Link | Code |
| AAMAS 2025 | Multi-Agent Reinforcement Learning with Selective State-Space Models | Link | Code |
| ICML 2025 | A Large Recurrent Action Model: xLSTM enables Fast Inference for Robotics Tasks | Link | Code |
| AAAI 2025 | GLAM: Global-Local Variation Awareness in Mamba-based World Model | Link | Code |
| CVPR 2025 | FlowRAM: Grounding Flow Matching Policy with Region-Aware Mamba Framework for Robotic Manipulation | Link |
Graph Learning
| Venue | Paper | Link | Code |
|---|---|---|---|
| KDD 2024 | Graph Mamba: Towards Learning on Graphs with State Space Models | Link | Code |
| KDD 2024 Workshop | Identifying Subphenotypes for Sepsis with Acute Kidney Injury via Multimodal Graph State Space Models | Link | |
| AAAI 2025 | DG-Mamba: Robust and Efficient Dynamic Graph Structure Learning with Selective State Space Models | Link | Code |
| AAAI 2025 | MOL-Mamba: Enhancing Molecular Representation with Structural & Electronic Insights | Link | Code |
| AAAI 2025 | BrainMAP: Learning Multiple Activation Pathways in Brain Networks | Link | Code |
| TMLR 2025 | DyGMamba: Efficiently Modeling Long-Term Temporal Dependency on Continuous-Time Dynamic Graphs with State Space Models | Link | |
| NeurIPS 2025 | DyG-Mamba: Continuous State Space Modeling on Dynamic Graphs | Link | Code |
| IJCNN 2025 | Topological Deep Learning with State-Space Models: A Mamba Approach for Simplicial Complexes | Link | |
| IJCAI 2025 | Mamba-Based Graph Convolutional Networks: Tackling Over-smoothing in Graph Neural Networks | Link | |
| IJCAI 2025 | SourceDetMamba: A Graph-aware State Space Model for Source Detection in Sequential Hypergraphs | Link | |
| ECML PKDD 2025 | GLADMamba: Unsupervised Graph-Level Anomaly Detection Powered by Selective State Space Model | Link | Code |
| AAAI 2026 | Dual Mamba for Node-Specific Representation Learning: Tackling Over-Smoothing with Selective State Space Modeling | Link | Code |
Audio
| Venue | Paper | Link | Code |
|---|---|---|---|
| IEEE SPL 2024 | Multichannel Long-Term Streaming Neural Speech Enhancement for Static and Moving Speakers | Link | Code |
| IMWUT 2025 | TRAMBA: A Hybrid Transformer and Mamba Architecture for Practical Audio and Bone Conduction Speech Super Resolution and Enhancement on Mobile and Wearable Platforms | Link | |
| Interspeech 2024 | Audio Mamba: Selective State Spaces for Self-Supervised Audio Representations | Link | Code |
| Interspeech 2024 | RawBMamba: End-to-End Bidirectional State Space Model for Audio Deepfake Detection | Link | Code |
| Interspeech 2024 | Exploring the Capability of Mamba in Speech Applications | Link | |
| SLT 2024 Workshop | An Analysis of Linear Complexity Attention Substitutes with BEST-RQ | Link | |
| SLT 2024 | Speech-Mamba: Long-Context Speech Recognition with Selective State Spaces Models | Link | Code |
| ICASSP 2025 | Mamba-based Segmentation Model for Speaker Diarization | Link | Code |
| SLT 2024 | Mamba-based Decoder-Only Approach with Bidirectional Speech Modeling for Speech Recognition | Link | Code |
| LAMIR 2024 Workshop | AEROMamba: An efficient architecture for audio super-resolution using generative adversarial networks and state space models | Link | Code |
| Interspeech 2025 | MASV: Speaker Verification with Global and Local Context Mamba | Link | |
| Expert Systems 2024 | A barking emotion recognition method based on Mamba and Synchrosqueezing Short-Time Fourier Transform | Link | Code |
| ICASSP 2025 | Mamba-SEUNet: Mamba UNet for Monaural Speech Enhancement | Link | |
| ICASSP 2025 | Temporal-Frequency State Space Duality: An Efficient Paradigm for Speech Emotion Recognition | Link | |
| APSIPA ASC 2024 | U-Mamba-Net: A highly efficient Mamba-based U-net style network for noisy and reverberant speech separation | Link | |
| AAAI 2025 | BSDB-Net: Band-Split Dual-Branch Network with Selective State Spaces Mechanism for Monaural Speech Enhancement | Link | |
| ICASSP 2025 | Improved Feature Extraction Network for Neuro-Oriented Target Speaker Extraction | Link | |
| IEEE SLT 2024 | An Investigation of Incorporating Mamba for Speech Enhancement | Link | |
| IEEE TASLP 2025 | Mamba in Speech: Towards an Alternative to Self-Attention | Link | |
| IEEE SLT 2024 | SSAMBA: Self-Supervised Audio Representation Learning with Mamba State Space Model | Link | Code |
| ICASSP 2025 | Speech Slytherin: Examining the Performance and Efficiency of Mamba for Speech Separation | Link | Code |
| CIAC 2025 | SELD-Mamba: Selective State-Space Model for Sound Event Localization and Detection with Source Distance Estimation | Link | |
| ICASSP 2025 | MusicMamba: A Dual-Feature Modeling Approach for Generating Chinese Traditional Music with SSM | Link | |
| ICASSP 2025 | Cross-attention Inspired Selective State Space Models for Target Sound Extraction | Link | |
| Interspeech 2025 | TF-Mamba: A Time-Frequency Network for Sound Source Localization | Link | |
| ICASSP 2025 | Vector Quantized Diffusion Model Based Speech Bandwidth Extension | Link | |
| ICASSP 2025 | Rethinking Mamba in Speech Processing by Self-Supervised Models | Link | |
| ICASSP 2025 | MambaFoley: Foley Sound Generation using Selective State-Space Models | Link | |
| ICASSP 2025 | Wave-U-Mamba: An End-To-End Framework For High-Quality And Efficient Speech Super Resolution | Link | |
| ICASSP 2025 | Self-supervised Learning for Acoustic Few-Shot Classification | Link | |
| ICASSP 2025 | Ultra-Low Latency Speech Enhancement - A Comprehensive Study | Link | |
| ICASSP 2025 | Leveraging Joint Spectral and Spatial Learning with MAMBA for Multichannel Speech Enhancement | Link | |
| ICASSP 2025 Oral | DeFT-Mamba: Universal Multichannel Sound Separation and Polyphonic Audio Classification | Link | |
| ICASSP 2025 | Mamba for Streaming ASR Combined with Unimodal Aggregation | Link | |
| ICLR 2025 | Joint Fine-tuning and Conversion of Pretrained Speech and Language Models towards Linear Complexity | Link | Code |
| ISCAS 2025 | CleanUMamba: A Compact Mamba Network for Speech Denoising using Channel Pruning | Link | |
| ICASSP 2025 | SepMamba: State-space models for speaker separation using Mamba | Link | Code |
| IEEE JSTSP 2025 | SAV-SE: Scene-aware Audio-Visual Speech Enhancement with Selective State Space Model | Link | |
| IEEE SPL 2025 | XLSR-Mamba: A Dual-Column Bidirectional State Space Model for Spoofing Attack Detection | Link | |
| ICASSP 2025 | BEST-STD: Bidirectional Mamba-Enhanced Speech Tokenization for Spoken Term Detection | Link | |
| ICASSP 2025 | TAME: Temporal Audio-based Mamba for Enhanced Drone Trajectory Estimation and Classification | Link | |
| Interspeech 2025 | xLSTM-SENet: xLSTM for Single-Channel Speech Enhancement | Link | Code |
| ICASSP 2025 | MSECG: Incorporating Mamba for Robust and Efficient ECG Super-Resolution | Link | |
| Interspeech 2025 | Leveraging Mamba with Full-Face Vision for Audio-Visual Speech Enhancement | Link | |
| Interspeech 2025 | BiCrossMamba-ST: Speech Deepfake Detection with Bidirectional Mamba Spectro-Temporal Cross-Attention | Link | |
| Interspeech 2025 | Universal Speech Enhancement with Regression and Generative Mamba | Link | |
| Interspeech 2025 | PARROT: Synergizing Mamba and Attention-based SSL Pre-Trained Models via Parallel Branch Hadamard Optimal Transport for Speech Emotion Recognition | Link | |
| Interspeech 2025 | Non-Intrusive Binaural Speech Intelligibility Prediction Using Mamba for Hearing-Impaired Listeners | Link | |
| ICASSP 2026 | BeatMamba: Bidirectional Selective State-Space Modeling for Efficient Beat Tracking | Link | |
| ICASSP 2026 | CMSA-Mamba: Hierarchical State Space Modeling for Audio-Based Depression Detection | Link |
Time Series
| Venue | Paper | Link | Code |
|---|---|---|---|
| ECAI 2024 | TimeMachine: A Time Series is Worth 4 Mambas for Long-term Forecasting | Link | Code |
| Information Fusion 2025 | TSCMamba: Mamba Meets Multi-View Learning for Time Series Classification | Link | |
| NeurIPS 2024 | Chimera: Effectively Modeling Multivariate Time Series with 2-Dimensional State Space Models | Link | |
| CIKM 2025 | SST: Multi-Scale Hybrid Mamba-Transformer Experts for Time Series Forecasting | Link | Code |
| IJCAI 2024 Workshop | SpoT-Mamba: Learning Long-Range Dependency on Spatio-Temporal Graphs with Selective State Spaces | Link | Code |
| ICECCE 2024 | Integration of Mamba and Transformer -- MAT for Long-Short Range Time Series Forecasting with Application to Weather Dynamics | Link | |
| IEEE IOTJ 2024 | HARMamba: Efficient and Lightweight Wearable Sensor Human Activity Recognition Based on Bidirectional Mamba | Link | |
| SLT 2024 | SWIM: Short-Window CNN Integrated with Mamba for EEG-Based Auditory Spatial Attention Decoding | Link | Code |
| IEEE GRSL 2024 | SPPMamba: State Space Models for Seismic Phase Arrival Picking | Link | |
| NeurIPS 2024 Workshop | Sequential Order-Robust Mamba for Time Series Forecasting | Link | Code |
| Scientific Reports 2024 | Application of multi-modal temporal neural network based on enhanced sparrow optimization in lithium battery life prediction | Link | |
| Scientific Reports 2024 | Mastering seismic time series response predictions using an attention-Mamba transformer model for bridge bearings and piers across varied testing conditions | Link | |
| ICASSP 2025 | SSM2Mel: State Space Model to Reconstruct Mel Spectrogram from the EEG | Link | |
| CGI 2024 | Mamba-Spike: Enhancing the Mamba Architecture with a Spiking Front-End for Efficient Temporal Data Processing | Link | Code |
| KDD 2025 | SDE: A Simplified and Disentangled Dependency Encoding Framework for State Space Models in Time Series Forecasting | Link | |
| ISCAS 2025 | SlimSeiz: Efficient Channel-Adaptive Seizure Prediction Using a Mamba-Enhanced Network | Link | Code |
| KDD 2025 | SSD-TS: Exploring the Potential of Linear State Space Models for Diffusion Models in Time Series Imputation | Link | |
| ICLR 2025 | FACTS: A Factored State-Space Framework For World Modelling | Link | Code |
| ICASSP 2025 | MSEMG: Surface Electromyography Denoising with a Mamba-based Efficient Network | Link | Code |
| ICBC 2025 | CryptoMamba: Leveraging State Space Models for Accurate Bitcoin Price Prediction | Link | Code |
| MICCAI 2025 | MambaMER: Adaptive EEG-Guided Multimodal Emotion Recognition with Mamba | Link | |
| ICLR 2025 Oral | High-Dynamic Radar Sequence Prediction for Weather Nowcasting Using Spatiotemporal Coherent Gaussian Representation | Link | |
| NeurIPS 2025 | RiverMamba: A State Space Model for Global River Discharge and Flood Forecasting | Link |
If you find this repository is useful for you, please cite our paper:
@misc{2024visual_mamba,
title={Visual Mamba: A Survey and New Outlooks},
author={Rui Xu and Shu Yang and Yihui Wang and Yu Cai and Bo Du and Hao Chen},
year={2024},
eprint={2404.18861},
archivePrefix={arXiv},
primaryClass={cs.CV}
}
Other works of HKUST SMART Lab:
@inproceedings{MambaMIL,
author = {Shu Yang and Yihui Wang and Hao Chen},
title = {MambaMIL: Enhancing Long Sequence Modeling with Sequence Reordering in Computational Pathology},
booktitle = {International Conference on Medical Image Computing and Computer-Assisted Intervention (MICCAI)},
volume = {15004},
pages = {296--306},
publisher = {Springer},
year = {2024}
}