SlowFastFormer for 3D human pose estimation
文献类型:期刊论文
作者 | Zhou Lu4![]() ![]() |
刊名 | Computer Vision and Image Understanding
![]() |
出版日期 | 2024 |
期号 | 243页码:103992 |
英文摘要 | 3D human pose estimation in videos aims at locating the human joints in the 3D space given a temporal
sequence. Motion information and skeleton context are two significant elements for pose estimation in videos.
In this paper, we propose a SlowFastFormer (slow-fast transformer) network where two branches with different
input rates are composed to encode these two different kinds of context. For the slow branch, skeleton context
is well learned at a higher frame rate. For the fast branch, motion information is captured at a lower frame rate.
Through these two branches, different kinds of context are encoded separately. We fuse these two branches
at a later stage to fully utilize the skeleton context and motion information. Afterwards, a blending module is
developed to promote the message exchange among multiple branches. In the blending stage, different kinds of
context information are exchanged and feature representation is enhanced consequently. Lastly, a hierarchical
supervision scheme is tailored where predictions of different levels are inferred in a progressive manner. Our
approach achieves competitive performance with lower computation complexity on several benchmarks, i.e.,
Human3.6M, MPI-INF-3DHP and HumanEva-I. |
语种 | 英语 |
源URL | [http://ir.ia.ac.cn/handle/173211/57150] ![]() |
专题 | 紫东太初大模型研究中心 |
通讯作者 | Chen Yingying |
作者单位 | 1.School of Artificial Intelligence, University of Chinese Academy of Sciences 2.Wuhan AI Research 3.Peng Cheng Laboratory 4.Foundation Model Research Center, Institute of Automation, Chinese Academy of Sciences |
推荐引用方式 GB/T 7714 | Zhou Lu,Chen Yingying,Wang Jinqiao. SlowFastFormer for 3D human pose estimation[J]. Computer Vision and Image Understanding,2024(243):103992. |
APA | Zhou Lu,Chen Yingying,&Wang Jinqiao.(2024).SlowFastFormer for 3D human pose estimation.Computer Vision and Image Understanding(243),103992. |
MLA | Zhou Lu,et al."SlowFastFormer for 3D human pose estimation".Computer Vision and Image Understanding .243(2024):103992. |
入库方式: OAI收割
来源:自动化研究所
浏览0
下载0
收藏0
其他版本
除非特别说明,本系统中所有内容都受版权保护,并保留所有权利。