科研速览 · Science Skim继续刷下去 · Keep skimming →
◆ Sensors (Basel, Switzerland)2026-08-01

BiasFormer: Structured Posterior-Guided Transformer for Temporal Action Segmentation.

Dongyue Zhou, Zhihui Shi, Haixia Wang, Hanqing Yang, Mingyue Yang, Chongchong Yu

原始摘要(英文原文)· Original abstract
Transformer-based architectures have become a dominant paradigm for temporal action segmentation because of their ability to capture long-range temporal dependencies. However, content-driven self-attention does not explicitly model action durations or feasible action transitions, which can lead to short spurious fragments and ambiguous action boundaries. To address this limitation, we propose BiasFormer, a unified end-to-end framework that couples feature-level structural guidance with decoder-level structural constraints. Specifically, frame-level confidence biases derived from structured Markov posteriors guide reliability-aware feature reweighting in frame-to-action cross-attention. We further introduce the unified segmental Markov head, which formulates temporal action segmentation as a segmental conditional random field with explicit duration modeling and transition support constraints. The structured objective is jointly optimized with the backbone, enabling structural information to influence both feature interaction and decoding. Extensive experiments on Breakfast, GTEA, EgoProceL, and EPIC-KITCHENS demonstrate competitive performance and improvements over FACT on most segment-level metrics. These results support the effectiveness of coupling posterior-guided feature refinement with structured decoding for improving segment-level temporal coherence.
读原文 · Read the paper ↗

AI 追问PRO

登录后使用 AI 追问

讨论区

登录后参与讨论

相关论文 · Related

BiasFormer: Structured Posterior-Guided Transformer for Temporal Action Segmentation. — 科研速览 Science Skim