跳到主要导航 跳到搜索 跳到主要内容

Multi-view Intention Recognition in Face-to-Face Communication

  • Pukun Chen
  • , Dongdong Weng*
  • , Xiaonuo Dongye
  • *此作品的通讯作者
  • Beijing Institute of Technology

科研成果: 书/报告/会议事项章节会议稿件同行评审

摘要

In this paper, we propose an intention recognition method based on generative dataset. Addressing the lack of intention datasets and the recognition methods in face-to-face communication scenarios, we analyze the motions corresponding to intentions and generate a motion dataset using a diffusion model. We then employ a Transformer-based method to map video to intention. In addition, we introduce a joint intention processing method that effectively handles the differences in motion semantics across different camera views, resulting in more accurate recognition outcomes in the case of multi-view data. Overall, this article summarizes a unified framework from acquisition to recognition.

源语言英语
主期刊名Image and Graphics Technologies and Applications - 19th Chinese Conference, IGTA 2024, Revised Selected Papers
编辑Yongtian Wang, Hua Huang
出版商Springer Science and Business Media Deutschland GmbH
327-338
页数12
ISBN(印刷版)9789819799183
DOI
出版状态已出版 - 2025
活动19th Chinese Conference on Image and Graphics Technologies and Applications, IGTA 2024 - Beijing, 中国
期限: 16 8月 202418 8月 2024

出版系列

姓名Communications in Computer and Information Science
2302 CCIS
ISSN(印刷版)1865-0929
ISSN(电子版)1865-0937

会议

会议19th Chinese Conference on Image and Graphics Technologies and Applications, IGTA 2024
国家/地区中国
Beijing
时期16/08/2418/08/24

指纹

探究 'Multi-view Intention Recognition in Face-to-Face Communication' 的科研主题。它们共同构成独一无二的指纹。

引用此