2021 IEEE International Conference on Acoustics, Speech and Signal Processing

6-11 June 2021 • Toronto, Ontario, Canada

Extracting Knowledge from Information

IEEE Signal Processing Society

Institute of Electrical and Electronics Engineers (IEEE)

2021 IEEE International Conference on Acoustics, Speech and Signal Processing

6-11 June 2021 • Toronto, Ontario, Canada

Extracting Knowledge from Information

Technical Program

Paper Detail

Paper ID	IVMSP-33.4
Paper Title	AN IMPROVED DEEP RELATION NETWORK FOR ACTION RECOGNITION IN STILL IMAGES
Authors	Wei Wu, Jiale Yu, Inner Mongolia University, China
Session	IVMSP-33: Action Recognition
Location	Gather.Town
Session Time:	Friday, 11 June, 14:00 - 14:45
Presentation Time:	Friday, 11 June, 14:00 - 14:45
Presentation	Poster
Topic	Image, Video, and Multidimensional Signal Processing: [IVSMR] Image & Video Sensing, Modeling, and Representation
IEEE Xplore Open Preview	Click here to view in IEEE Xplore
Virtual Presentation	Click here to watch in the Virtual Conference
Abstract	Contextual information has been widely utilized in visual recognition tasks. This is especially true for action recognition, because contextual information such as objects interacting with human and the scene where the action is performed is inseparable from action categories. To this end, we propose an efficient relation module that combines Human-Object and Scene-Object relations for action recognition. Specifically, Human-Object interaction submodule can capture more accurate appearance and spatial relation to build human-object interaction pairs. And Scene-Object interaction submodule can learn the probability of the objects involved in the scene to help discover the key interaction pair. We conduct extensive experiments on Stanford 40 and Pascal Voc 2012 Action datasets to verify our model, and experimental results show that our method achieves superior performance on these two datasets. Especially, we gain the best results on the Stanford 40 dataset compared with state-of-the-arts.