科研速览 · Science Skim继续刷下去 · Keep skimming →
◆ Lecture notes in computer science2026-08-02· Computer science

Semantically Stable Image Composition Analysis via Saliency and Gradient Vector Flow Fusion

Armin Dadras, Robert Sablatnig, Franziska Brückner, Markus Seidl

原始摘要(英文原文)· Original abstract
Abstract The reliable computational assessment of photographic composition requires features that are discriminative of spatial layout yet robust to semantic content. This paper proposes a low-level representation grounded in the assumption that composition can be understood as the flow of visual attention across geometric structure. We introduce VFCNet, which fuses saliency and edge information into a gradient vector flow (GVF) field. The model computes dual-stream GVF representations, integrates them via attention, and extracts multi-scale flow features with a DINOv3 backbone. VFCNet achieves state-of-the-art performance on the PICD benchmark (CDA-1: 0.683, CDA-2: 0.629), improving by 33.1% and 36.1% over the previous best method. We also show that a simple classifier on self-supervised DINOv3 features substantially outperforms more sophisticated, composition-specialized models. Code is available at https://github.com/ADadras/VFCNet .
读原文 · Read the paper ↗

AI 追问PRO

登录后使用 AI 追问

讨论区

登录后参与讨论

相关论文 · Related

Semantically Stable Image Composition Analysis via Saliency and Gradient Vector Flow Fusion — 科研速览 Science Skim