REVIEW 3 cited by
Computation-Efficient Era: A Comprehensive Survey of State Space Models in Medical Image Analysis
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
abstract
Sequence modeling plays a vital role across various domains, with recurrent neural networks being historically the predominant method of performing these tasks. However, the emergence of transformers has altered this paradigm due to their superior performance. Built upon these advances, transformers have conjoined CNNs as two leading foundational models for learning visual representations. However, transformers are hindered by the $\mathcal{O}(N^2)$ complexity of their attention mechanisms, while CNNs lack global receptive fields and dynamic weight allocation. State Space Models (SSMs), specifically the \textit{\textbf{Mamba}} model with selection mechanisms and hardware-aware architecture, have garnered immense interest lately in sequential modeling and visual representation learning, challenging the dominance of transformers by providing infinite context lengths and offering substantial efficiency maintaining linear complexity in the input sequence. Capitalizing on the advances in computer vision, medical imaging has heralded a new epoch with Mamba models. Intending to help researchers navigate the surge, this survey seeks to offer an encyclopedic review of Mamba models in medical imaging. Specifically, we start with a comprehensive theoretical review forming the basis of SSMs, including Mamba architecture and its alternatives for sequence modeling paradigms in this context. Next, we offer a structured classification of Mamba models in the medical field and introduce a diverse categorization scheme based on their application, imaging modalities, and targeted organs. Finally, we summarize key challenges, discuss different future research directions of the SSMs in the medical domain, and propose several directions to fulfill the demands of this field. In addition, we have compiled the studies discussed in this paper along with their open-source implementations on our GitHub repository.
Forward citations
Cited by 3 Pith papers
-
MambaU-Lite: A Lightweight Model based on Mamba and Integrated Channel-Spatial Attention for Skin Lesion Segmentation
MambaU-Lite, a 0.42M-parameter hybrid Mamba-CNN model, reports DSC/IoU of 0.9057/0.8361 on ISIC2018 and 0.9572/0.9189 on PH2, best among the compared lightweight models.
-
Medical Image Segmentation Using Advanced Unet: VMSE-Unet and VM-Unet CBAM+
Adding Squeeze-and-Excitation and CBAM attention to VM-UNet is reported to improve segmentation metrics, but the paper's own data contradict the claim that VMSE-Unet wins on all metrics.
-
A Study on the Performance of U-Net Modifications in Retroperitoneal Tumor Segmentation
ViLU-Net, a U-Net built with Vision LSTM blocks, reports the best segmentation scores on a new retroperitoneal tumor CT dataset and on FLARE22, but the comparison lacks error bars and omits the closest prior architecture.
Discussion (0). Continue with ORCID to comment.