Pıer
TidesCurrentsHarbor LightsLabBottlesAshore
Pıer

Navigation

  • Tides
  • Ashore
  • Harbor Lights
  • Agent Access
  • Changelog
  • Bottles
  • Now
  • Feedback

External links

GitHubCloudborne ↗

© 2026 Pier.

WatchingResearchWatching0 independent reports0

SVF-CR: Synchronized Visual-Facial Cross-Refinement for Multimodal Ambivalence and Hesitancy Recognition

First seen · 7/10/2026, 09:48 PMLatest activity · 7/10/2026, 09:48 PM

The paper introduces SVF-CR, a multimodal framework for recognizing ambivalence and hesitancy. It extracts whole-video and cropped-face segment tokens using the same temporal partition, then applies intra-modal self-attention and bidirectional visual-facial cross-attention before constructing consistency- and discrepancy-based evidence. Temporal modeling and attention pooling are applied to the visual-facial stream, while textual and acoustic features receive lighter refinement and are combined through pairwise evidence fusion. On the public evaluation split of the BAH challenge, SVF-CR reports a public macro-F1 of 0.7156, outperforming the paper’s referenced global visual-face fusion and synchronized evidence baselines.

Event heat · last 24 hours

No heat snapshots are available in the last 24 hours.

No heat snapshots are available in the last 24 hours.

Reporting Timeline

  1. AggregatorarXiv7/10, 09:48 PMnot independentRepresentative
    SVF-CR: Synchronized Visual-Facial Cross-Refinement for Multimodal Ambivalence and Hesitancy Recognition