A Semantic and Motion-Aware Spatiotemporal Transformer Network for Action Detection

This paper presents a novel spatiotemporal transformer network that introduces several original components to detect actions in untrimmed videos. First, the multi-feature selective semantic attention model calculates the correlations between spatial and motion features to model spatiotemporal intera...

Description complète

Détails bibliographiques
Publié dans:IEEE transactions on pattern analysis and machine intelligence. - 1979. - 46(2024), 9 vom: 14. Sept., Seite 6055-6069
Auteur principal: Korban, Matthew (Auteur)
Autres auteurs: Youngs, Peter, Acton, Scott T
Format: Article en ligne
Langue:English
Publié: 2024
Accès à la collection:IEEE transactions on pattern analysis and machine intelligence
Sujets:Journal Article