摘要
Detecting fake news has become increasingly challenging in the era of multimodal social media, where deceptive content often combines misleading text with incongruent images. Existing methods frequently suffer from three key limitations: (1) misaligned semantic representations across modalities, (2) reliance on overly complex or inefficient fusion mechanisms, and (3) often overlook cross-modal semantic (in)consistencies, which are widely regarded as critical cues for detecting fake news. To address these challenges, we propose the Cross-Modal Aligned Attention Network (CMAAN), an efficient framework tailored for multimodal rumor detection. CMAAN utilizes CLIP to obtain semantically aligned text and image embeddings, which are then integrated into a compact joint representation via a lightweight gating-and-weighting module. Central to our approach is a dual-stage cross-modal attention mechanism: the first stage refines unimodal features under the guidance of the fused representation, while the second facilitates bidirectional interactions to enable effective cross-modal reasoning. Experiments on two widely used real-world multimodal rumor detection datasets demonstrate the effectiveness of the proposed approach.
| 源语言 | 英语 |
|---|---|
| 页(从-至) | 593-598 |
| 页数 | 6 |
| 期刊 | Proceedings of the International Conference on Computer Supported Cooperative Work in Design, CSCWD |
| 期 | 2026 |
| DOI | |
| 出版状态 | 已出版 - 2026 |
| 活动 | 29th International Conference on Computer Supported Cooperative Work in Design, CSCWD 2026 - Fuzhou, 中国 期限: 13 5月 2026 → 15 5月 2026 |
学术指纹
探究 'Dual-Stage Cross-Modal Attention Network for Multimodal Fake News Detection' 的科研主题。它们共同构成独一无二的学术指纹。引用此
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver