Human Actions

Recognizing, localizing, and detecting human actions in video, from fully-supervised to weakly- and semi-supervised settings.

(Kumar & Rawat, 2022) (Rana & Rawat, 2021) (Rana & Rawat, 2022) (Rana & Rawat, 2023) (Rana et al., 2025) (Dave et al., 2022) (Rizve et al., 2020) (Tirupattur et al., 2021) (Tirupattur et al., 2021) (Kerrigan et al., 2021) (Ahmad et al., 2025) (Ahmad et al., 2023) (Singh et al., 2024) (Modi et al., 2023) (Grover et al., 2023) (Modi et al., 2022) (Sia & Rawat, 2025) (Abdullah et al., 2025) (Abdullah et al., 2026) (Swetha et al., 2021) (Demir et al., 2020) (Xu et al., 2022) (Kumar et al., 2025) (Dave et al., 2021)

References

2026

  1. vlm.jpg
    Learning to Deny: Action Denial in Multimodal Large Language Models
    Raiyaan Abdullah, Shehreen Azad, and Yogesh Singh Rawat
    In European Conference on Computer Vision (ECCV), 2026

2025

  1. video-action.jpg
    OmViD: Omni-supervised active learning for video action detection
    Aayush Rana, Akash Kumar, Vibhav Vineet, and Yogesh S Rawat
    In IEEE/CVF International Conference on Computer Vision Workshops (ICCVW), 2025
  2. T2L: Efficient Zero-Shot Action Recognition with Temporal Token Learning
    Shahzad Ahmad, Sukalpa Chanda, and Yogesh S. Rawat
    Transactions on Machine Learning Research (TMLR), 2025
  3. Scaling Open-Vocabulary Action Detection
    Zhen Hao Sia and Yogesh Singh Rawat
    In IEEE/CVF International Conference on Computer Vision Workshops (ICCVW), 2025
  4. Punching Bag vs. Punching Person: Motion Transferability in Videos
    Raiyaan Abdullah, Jared Claypoole, Michael Cogswell, Ajay Divakaran, and Yogesh Rawat
    In IEEE/CVF International Conference on Computer Vision (ICCV), 2025
  5. Stable Mean Teacher for Semi-supervised Video Action Detection
    Akash Kumar, Sirshapan Mitra, and Yogesh Singh Rawat
    In AAAI Conference on Artificial Intelligence, 2025

2024

  1. Semi-supervised active learning for video action detection
    Ayush Singh, Aayush J Rana, Akash Kumar, Shruti Vyas, and Yogesh Singh Rawat
    In Proceedings of the AAAI Conference on Artificial Intelligence, 2024

2023

  1. Hybrid active learning via deep clustering for video action detection
    Aayush J Rana and Yogesh S Rawat
    In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, 2023
  2. EZ-CLIP: Efficient Zero-shot Video Action Recognition
    Shahzad Ahmad, Sukalpa Chanda, and Yogesh S Rawat
    arXiv preprint arXiv:2312.08010, 2023
  3. On occlusions in video action detection: Benchmark datasets and training recipes
    Rajat Modi, Vibhav Vineet, and Yogesh Rawat
    Advances in Neural Information Processing Systems, 2023
  4. Revealing the unseen: Benchmarking video action recognition under occlusion
    Shresth Grover, Vibhav Vineet, and Yogesh Rawat
    Advances in Neural Information Processing Systems, 2023

2022

  1. End-to-End Semi-Supervised Learning for Video Action Detection
    Akash Kumar and Yogesh Singh Rawat
    In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, 2022
  2. Are all Frames Equal? Active Sparse Labeling for Video Action Detection
    Aayush Rana and Yogesh S Rawat
    In Advances in Neural Information Processing Systems, 2022
  3. surveillance.jpg
    GabriellaV2: Towards better generalization in surveillance videos for action detection
    Ishan Dave, Zacchaeus Scheffer, Akash Kumar, Sarah Shiraz, Yogesh Singh Rawat, and Mubarak Shah
    In IEEE/CVF Winter Conference on Applications of Computer Vision Workshops (WACVW), 2022
  4. Video Action Detection: Analysing Limitations and Challenges
    Rajat Modi, Aayush Jung Rana, Akash Kumar, Praveen Tirupattur, Shruti Vyas, Yogesh Singh Rawat, and Mubarak Shah
    In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition Workshops, 2022
  5. Don’t Pour Cereal into Coffee: Differentiable Temporal Logic for Temporal Action Segmentation
    Ziwei Xu, Yogesh S Rawat, Yongkang Wong, Mohan Kankanhalli, and Mubarak Shah
    In Advances in Neural Information Processing Systems, 2022

2021

  1. video-action.jpg
    We Don’t Need Thousand Proposals: Single Shot Actor-Action Detection in Videos
    Aayush J Rana and Yogesh S Rawat
    In IEEE/CVF Winter Conference on Applications of Computer Vision (WACV), 2021
  2. Modeling Multi-Label Action Dependencies for Temporal Action Localization
    Praveen Tirupattur, Kevin Duarte, Yogesh Rawat, and Mubarak Shah
    In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, 2021
  3. video-action.jpg
    Tinyaction challenge: Recognizing real-world low-resolution activities in videos
    Praveen Tirupattur, Aayush J Rana, Tushar Sangam, Shruti Vyas, Yogesh S Rawat, and Mubarak Shah
    arXiv preprint arXiv:2107.11494, 2021
  4. Reformulating zero-shot action recognition for multi-label actions
    Alec Kerrigan, Kevin Duarte, Yogesh Rawat, and Mubarak Shah
    Advances in Neural Information Processing Systems, 2021
  5. video-action.jpg
    Unsupervised Discriminative Embedding for Sub-Action Learning in Complex Activities
    Sirnam Swetha, Hilde Kuehne, Yogesh S Rawat, and Mubarak Shah
    In 2021 IEEE International Conference on Image Processing, 2021
  6. video-action.jpg
    "Knights": First Place Submission for VIPriors21 Action Recognition Challenge at ICCV 2021
    Ishan Dave, Naman Biyani, Brandon Clark, Rohit Gupta, Yogesh Rawat, and Mubarak Shah
    arXiv preprint arXiv:2110.07758, 2021

2020

  1. surveillance.jpg
    Gabriella: An Online System for Real-Time Activity Detection in Untrimmed Security Videos
    Mamshad Nayeem Rizve, Ugur Demir, Praveen Tirupattur, Aayush Jung Rana, Kevin Duarte, Ishan Dave, Yogesh Singh Rawat, and Mubarak Shah
    In 25th International Conference on Pattern Recognition (ICPR), 2020
  2. video-action.jpg
    TinyVIRAT: Low-resolution Video Action Recognition
    Ugur Demir, Yogesh S Rawat, and Mubarak Shah
    In 25th International Conference on Pattern Recognition (ICPR), 2020