Skeletal Keypoint-Based Transformer Model for Human Action Recognition in Aerial Videos
Several efforts have been made to develop effective and robust vision-based solutions for human action recognition in aerial videos. Generally, the existing methods rely on the extraction of either spatial features (patch-based methods) or skeletal key points (pose-based methods) that are fed to a c...
Κύριοι συγγραφείς: | , , , , , |
---|---|
Μορφή: | Άρθρο |
Γλώσσα: | English |
Έκδοση: |
IEEE
2024-01-01
|
Σειρά: | IEEE Access |
Θέματα: | |
Διαθέσιμο Online: | https://ieeexplore.ieee.org/document/10400454/ |