A Semantic and Motion-Aware Spatiotemporal Transformer Network for Action Detection

This paper presents a novel spatiotemporal transformer network that introduces several original components to detect actions in untrimmed videos. First, the multi-feature selective semantic attention model calculates the correlations between spatial and motion features to model spatiotemporal intera...

Description complète

Détails bibliographiques
Publié dans:	IEEE transactions on pattern analysis and machine intelligence. - 1979. - 46(2024), 9 vom: 14. Sept., Seite 6055-6069
Auteur principal:	Korban, Matthew (Auteur)
Autres auteurs:	Youngs, Peter, Acton, Scott T
Format:	Article en ligne
Langue:	English
Publié:	2024
Accès à la collection:	IEEE transactions on pattern analysis and machine intelligence
Sujets:	Journal Article

Accès en ligne	Volltext