Skip to Main Content (Press Enter)

Logo UNIMORE
  • ×
  • Home
  • Corsi
  • Insegnamenti
  • Professioni
  • Persone
  • Pubblicazioni
  • Strutture
  • Terza Missione
  • Attività
  • Competenze

UNI-FIND
Logo UNIMORE

|

UNI-FIND

unimore.it
  • ×
  • Home
  • Corsi
  • Insegnamenti
  • Professioni
  • Persone
  • Pubblicazioni
  • Strutture
  • Terza Missione
  • Attività
  • Competenze
  1. Pubblicazioni

Video action detection by learning graph-based spatio-temporal interactions

Articolo
Data di Pubblicazione:
2021
Citazione:
Video action detection by learning graph-based spatio-temporal interactions / Tomei, M., Baraldi, L., Calderara, S., Bronzin, S., Cucchiara, R.. - In: COMPUTER VISION AND IMAGE UNDERSTANDING. - ISSN 1077-3142. - 206:(2021), pp. 1-9. [10.1016/j.cviu.2021.103187]
Abstract:
Action Detection is a complex task that aims to detect and classify human actions in video clips. Typically, it has been addressed by processing fine-grained features extracted from a video classification backbone. Recently, thanks to the robustness of object and people detectors, a deeper focus has been added on relationship modelling. Following this line, we propose a graph-based framework to learn high-level interactions between people and objects, in both space and time. In our formulation, spatio-temporal relationships are learned through self-attention on a multi-layer graph structure which can connect entities from consecutive clips, thus considering long-range spatial and temporal dependencies. The proposed module is backbone independent by design and does not require end-to-end training. Extensive experiments are conducted on the AVA dataset, where our model demonstrates state-of-the-art results and consistent improvements over baselines built with different backbones. Code is publicly available at https://github.com/aimagelab/STAGE_action_detection.
Tipologia CRIS:
Articolo su rivista
Keywords:
Action detection; Graph learning; Video understanding;
Elenco autori:
Tomei, Matteo; Baraldi, Lorenzo; Calderara, Simone; Bronzin, Simone; Cucchiara, Rita
Autori di Ateneo:
BARALDI LORENZO
CALDERARA Simone
CUCCHIARA Rita
Link alla scheda completa:
https://iris.unimore.it/handle/11380/1235540
Link al Full Text:
https://iris.unimore.it//retrieve/handle/11380/1235540/572305/1-s2.0-S107731422100031X-main.pdf
Pubblicato in:
COMPUTER VISION AND IMAGE UNDERSTANDING
Journal
  • Utilizzo dei cookie

Realizzato con VIVO | Designed by Cineca | 26.7.2.0