---
title: GroupMamba: Efficient Group-Based Visual State Space Model
url: https://www.ml-quant.com/papers/arxiv/2407.13772/
site: ML-Quant (https://www.ml-quant.com)
updated: 2026-09-26
license: Summaries CC BY 4.0; links go to the original sources
index: https://www.ml-quant.com/llms.txt
identifier: arXiv:2407.13772
source_url: https://arxiv.org/abs/2407.13772
featured: 2024-07-24
citations: 23
topic: Other
---


# GroupMamba: Efficient Group-Based Visual State Space Model

Efficient Visual State Model: The Modulated Group Mamba layer is introduced for state-space models, effectively addressing scaling issues in computer vision tasks and improving performance in image classification, object detection, and segmentation.

- Source: https://arxiv.org/abs/2407.13772
- Identifier: arXiv:2407.13772
- Released: 2024-07-18
- First featured: Quant Letter No. 58 (2024-07-24): https://www.ml-quant.com/issues/2024-07-24/
- Citations (Semantic Scholar): 23
- Published in: 2025 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)
- Topic: Other

## Related

- [Mamba-Reg: Vision Mamba Also Needs Registers](https://www.ml-quant.com/papers/arxiv/2405.14858/): The paper reveals artifacts in Vision Mamba's feature maps and introduces Mamba-R, a new architecture with register tokens that improves feature maps, performance, and scalability.
- [MambaStock: Selective state space model for stock prediction](https://www.ml-quant.com/papers/arxiv/2402.18959/): The paper presents MambaStock, a new Mamba-based model for predicting stock prices using historical market data, which outperforms previous methods in accuracy, aiding investors in making informed decisions.
- [Hierarchical State Space Models for Continuous Sequence-to-Sequence Modeling](https://www.ml-quant.com/papers/arxiv/2402.10211/): Hierarchical State-Space Models, a new method for continuous sequential prediction, outperforms existing models in predicting sequences from raw sensory data, showing efficient scaling to smaller datasets and compatibility with existing data-filtering techniques.
- [Depth Anything V2](https://www.ml-quant.com/papers/arxiv/2406.09414/): Depth Anything V2 is a new model for monocular depth estimation, using synthetic and large-scale pseudo-labeled real images for faster, more accurate results and setting a new evaluation benchmark.
- [MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark](https://www.ml-quant.com/papers/arxiv/2406.01574/): MMLU-Pro, an improved dataset, expands the Massive Multitask Language Understanding benchmark by adding tougher questions and more choices, serving as a better benchmark to monitor progress in the field.
- [Qwen2.5-Coder Technical Report](https://www.ml-quant.com/papers/arxiv/2409.12186/): The report unveils the Qwen2.5-Coder series, an improvement from its predecessor, showcasing remarkable code generation abilities and achieving top-tier performance in various code-related tasks.
