PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
March 29, 2026Scientific Reports0 citationsOpen Access

ScaleMamba-YOLO: a multi-scale MambaYOLO for medical object detection

XQXiao QinQQQuanmei QianXLXiaosen Li

Key Points

  • To develop an advanced framework for automated lesion detection that addresses scale variability and background interference.
  • Introduced ScaleMamba-YOLO framework
  • Implemented Medical Multi-scale Local Feature Enhancement Block (MMLFE-Block)
  • Utilized heterogeneous convolutional kernels for hierarchical perception
  • Designed Partial-Enhanced C2F (PEC2F) module for selective feature processing
  • Achieved Average Precision scores of 72.7%, 65.0%, 85.7%, and 64.6% across datasets
  • Consistent improvements of 1.7% to 2.3% over standard MambaYOLO baseline

Abstract

The efficacy of automated lesion detection in clinical settings is often hampered by two primary factors: the vast range of pathological scales and the presence of non-target anatomical interference. Standard Mamba-based detectors, while efficient, frequently suffer from fixed receptive fields and background signal leakage. To resolve these challenges, we introduce ScaleMamba-YOLO, an enhanced medical object detection framework that integrates selective state-space modeling with adaptive local feature refinement. Our approach features two innovative structural components: first, a Medical Multi-scale Local Feature Enhancement Block (MMLFE-Block) is positioned at the frontend to diversify the initial receptive field. By utilizing a parallel architecture with heterogeneous convolutional kernels (11, 33, and 55), the model achieves comprehensive hierarchical perception, enabling the concurrent identification of minute calcified spots and extensive diffuse lesions. Second, a Partial-Enhanced C2F (PEC2F) module is designed to refine feature aggregation post-global modeling. This component employs partial convolution (PConv) to selectively process salient feature channels, effectively filtering out irrelevant background noise from normal tissue structures. The robust performance of ScaleMamba-YOLO was validated across three specialized medical datasets (Br35H, BCCD, and PLoPy) and a standard scene dataset (VOC0712). The model recorded Average Precision (AP) scores of 72. 7%, 65. 0%, 85. 7%, and 64. 6%, respectively. These metrics represent consistent improvements of 1. 7% to 2. 3% over the MambaYOLO baseline, underscoring the system’s potential for high-fidelity diagnostic assistance in real-time clinical workflows.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Qin et al. (2026) studied this question.

synapsesocial.com/papers/69c8c28cde0f0f753b39cef9https://doi.org/10.1038/s41598-026-37258-8
Ask AI
Helpful
Bookmark
Share
View Full Paper