PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
February 21, 2026Sensors0 citationsOpen Access

FreqPose: Frequency-Aware Diffusion with Fractional Gabor Filters and Global Pose–Semantic Alignment

View Full Paper
MWMengdi WangBWBing WangHCHuiling Chen

Key Points

  • The aim is to improve pose-guided person image generation by addressing texture detail loss and semantic identity consistency during pose changes.
  • Developed a frequency-aware diffusion framework incorporating fractional-order Gabor filters.
  • Implemented a global semantic-pose alignment module using cross-modal attention mechanisms.
  • Conducted experiments on the DeepFashion and Market1501 datasets to evaluate the performance.
  • Achieved higher structural integrity and natural texture in generated images.
  • Outperformed existing methods in SSIM (Structural Similarity Index), FID (Fréchet Inception Distance), and perceptual quality metrics.
  • Effectively maintained high-frequency details like hair strands and fabric wrinkles.

Abstract

The task of pose-guided person image generation has long been confronted with two major challenges: high-frequency texture details tend to blur and be lost during appearance transfer, while the semantic identity of the person is difficult to maintain consistently during pose changes. To address these issues, this paper proposes a diffusion-based generative framework that integrates frequency awareness and global semantic alignment. The framework consists of two core modules: a multi-level fractional-order Gabor frequency-aware network, which accurately extracts and reconstructs high-frequency texture features such as hair strands and fabric wrinkles, enhances image detail fidelity through fractional-order filtering and complex domain modeling; and a global semantic-pose alignment module that utilizes a cross-modal attention mechanism to establish a global mapping between pose features and appearance semantics, ensuring pose-driven semantic alignment and appearance consistency. The collaborative function of these two modules ensures that the generated results maintain structural integrity and natural textures even under complex pose variations and large-angle rotations. The experimental results on the DeepFashion and Market1501 datasets demonstrate that the proposed method outperforms existing state-of-the-art approaches in terms of SSIM, FID, and perceptual quality, validating the effectiveness of the model in enhancing texture fidelity and semantic consistency.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Wang et al. (2026) studied this question.

synapsesocial.com/papers/69994c80873532290d020fd2https://doi.org/10.3390/s26041334
Ask AI
Helpful
Bookmark
Share
View Full Paper