中国科学院机构知识库网格系统: Learning Scene-Aware Spatio-Temporal GNNs for Few-Shot Early Action Prediction

中国科学院机构知识库网格

Chinese Academy of Sciences Institutional Repositories Grid

Learning Scene-Aware Spatio-Temporal GNNs for Few-Shot Early Action Prediction

文献类型：期刊论文


作者	Hu, Yufan 1,2; Gao, Junyu3,4 ; Xu, Changsheng1,3,4
刊名	IEEE Transactions on Multimedia
出版日期	2022
页码	1-14
英文摘要	We aim to address a new task named few-shot early action prediction (FS-EAP) that learns classifiers for novel actions from only a few partially observed videos. We argue that the task is extremely challenging since the partially observed videos do not contain enough action information in a few-shot environment. To tackle this task, in this paper, we propose a scene-aware spatio-temporal graph neural network (SA-STGNN) by leveraging the fine-grained spatio-temporal interactions in the video scenes. Specifically, we first generate a spatio-temporal graph corresponding to the partially observed video to capture comprehensive spatio-temporal correlations. Then we utilize the spatio-temporal graph as the input of our SA-STGNN and predict the augmented video features corresponding to the complete video. The architecture uses several scene-aware learning blocks, which are a combination of edge fusion graph neural layers and temporal gated convolutional layers to jointly model spatial and temporal dependencies. Finally, we employ an early action predictor to exploit the learned video features for predicting actions in the few-shot setting. Extensive experimental results on two widely adopted video datasets demonstrate the effectiveness of our approach and its superior performance over the state-of-the-art approaches.
源URL	[http://ir.ia.ac.cn/handle/173211/51527]
专题	多模态人工智能系统全国重点实验室
作者单位	1.Peng Cheng Laboratory 2.Hefei University of Technology 3.National Lab of Pattern Recognition, Institute of Automation, Chinese Academy of Sciences 4.School of Artifical Intelligence, University of Chinese Academy of Sciences
推荐引用方式 GB/T 7714	Hu, Yufan,Gao, Junyu,Xu, Changsheng. Learning Scene-Aware Spatio-Temporal GNNs for Few-Shot Early Action Prediction[J]. IEEE Transactions on Multimedia,2022:1-14.
APA	Hu, Yufan,Gao, Junyu,&Xu, Changsheng.(2022).Learning Scene-Aware Spatio-Temporal GNNs for Few-Shot Early Action Prediction.IEEE Transactions on Multimedia,1-14.
MLA	Hu, Yufan,et al."Learning Scene-Aware Spatio-Temporal GNNs for Few-Shot Early Action Prediction".IEEE Transactions on Multimedia (2022):1-14.

入库方式： OAI收割

来源：自动化研究所

浏览0

下载0

收藏0

其他版本

除非特别说明，本系统中所有内容都受版权保护，并保留所有权利。