科研速览 · Science Skim继续刷下去 · Keep skimming →
◆ Journal of King Saud University - Computer and Information Sciences2025-12-13· Computer science

Frequency-prior enhanced network for facial expression recognition via dynamic large kernels and dual-domain learning

Chuanyu Cai, Ke Chen

原始摘要(英文原文)· Original abstract
Facial Expression Recognition in natural scenes constitutes a critical research direction in affective computing. Prevailing approaches predominantly rely on RGB spatial domain modeling, neglecting the inherent texture prior information in the frequency domain, which consequently limits their adaptability to challenging scenarios. Furthermore, conventional methods typically depend on single-domain spatial features, failing to effectively exploit the complementary characteristics between spatial and frequency domains, thereby limiting the model’s capacity for subtle expression representation. To address these limitations, this paper proposes FPNet, a dual-domain collaborative framework for FER. Specifically, we design three core components: (1) DyLKBlock constructs dynamic spatial feature extraction through cascaded large-kernel convolutions, achieving an equivalent 23 \(\times \) 23 receptive fields while balancing global modeling capability with linear computational complexity; (2) FPBlock utilizes D iscrete W avelet T ransform (DWT) to achieve lossless downsampling, preserving multi-scale texture details; (3) CAFusion facilitates bidirectional cross-domain feature interaction through an attention mechanism, ensuring the preservation of discriminative frequency domain characteristics in cross-domain data. Extensive experiments on three benchmark datasets demonstrate that FPNet significantly outperforms SOTA methods. Visualization analyses further validate that the frequency domain priors in FPNet effectively capture subtle expressions in complex scenarios, with precise focus on key facial regions.
读原文 · Read the paper ↗

AI 追问PRO

登录后使用 AI 追问

讨论区

登录后参与讨论

相关论文 · Related

Frequency-prior enhanced network for facial expression recognition via dynamic large kernels and dual-domain learning — 科研速览 Science Skim