PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
February 12, 20260 citationsOpen Access

Enhancing BlazeBlock with Convolutional Block Attention and self-attention on agricultural vision applications

ACArnab Ghosh ChowdhuryASAmos SmithMAMartin Atzmueller

Key Points

  • The main aim is to enhance the BlazeBlock CNN architecture with attention mechanisms for better agricultural vision performance.
  • Integration of Convolutional Block with Spatial Self-Attention Module (CBwSSAM) in BlazeBlock.
  • Utilization of Convolutional Block Attention Module (CBAM) for channel attention.
  • Replacement of spatial attention with MobileViT-based self-attention for a global representation.
  • Evaluation on six datasets for image classification and transfer learning, and three for semantic segmentation.
  • Significant improvements in image classification accuracy across all evaluated datasets.
  • Enhanced performance in semantic segmentation tasks compared to previous models.
  • Demonstrated capability of CBwSSAM to seamlessly integrate into various CNN architectures.

Abstract

Convolutional neural networks (CNNs) perform a pivotal role in agricultural vision applications. The CNN-based BlazeBlock has been previously proposed as a building block of lightweight BlazeFace models for face detection tasks on mobile GPUs. Thus, the utilization of BlazeBlock is often advantageous in order to offer mobile and edge computing vision services in precision farming. Moreover, the Convolutional Block Attention Module (CBAM) enhances the representational power of CNNs by emphasizing important features through the use of channel attention and spatial attention modules. Although CBAM uses a convolutional layer to generate a spatial attention map on the feature descriptor in spatial attention, it is spatially local in nature. Moreover, employing self-attention on the feature descriptor is beneficial to learn the global representation. Hence, a MobileViT-based self-attention module can substitute the prior spatial attention module. In this paper, we present an attention-based BlazeBlock that uses the Convolutional Block with Spatial Self-Attention Module (CBwSSAM) by harnessing the CBAM channel attention and a MobileViT-based spatial self-attention. CBwSSAM can be integrated into any CNN architecture seamlessly. We analyze the effectiveness of the attention-based BlazeBlock in image classification and semantic segmentation. Our evaluation results on six datasets for image classification and transfer learning, and three datasets for semantic segmentation exemplify the efficacy of the proposed enhanced BlazeBlock in such agricultural vision application contexts.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Chowdhury et al. (2026) studied this question.

synapsesocial.com/papers/698d6dc15be6419ac0d52f61https://doi.org/10.18420/giljt2026_04
Ask AI
Helpful
Bookmark
Share
View Full Paper