Video Understanding

Foundational video representation learning, cross-view and novel-view synthesis, and video segmentation.

References

2021

  1. Pose-guided Generative Adversarial Net for Novel View Action Synthesis
    Xianhang Li, Junhao Zhang, Kunchang Li, Shruti Vyas, and Yogesh S Rawat
    In Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision, 2021
  2. video-action.jpg
    Novel View Video Prediction Using a Dual Representation
    Sarah Shiraz, Krishna Regmi, Shruti Vyas, Yogesh S Rawat, and Mubarak Shah
    In IEEE International Conference on Image Processing, 2021
  3. LARNet: Latent Action Representation for Human Action Synthesis
    Naman Biyani, Aayush J Rana, Shruti Vyas, and Yogesh S Rawat
    In The British Machine Vision Conference (BMVC), 2021
  4. video-action.jpg
    Plm: Partial label masking for imbalanced multi-label classification
    Kevin Duarte, Yogesh Rawat, and Mubarak Shah
    In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, 2021

2020

  1. video-action.jpg
    Multi-view Action Recognition using Cross-view Video Prediction
    Shruti Vyas, Yogesh S Rawat, and Mubarak Shah
    In European Conference on Computer Vision (ECCV), 2020
  2. video-action.jpg
    A Recurrent Transformer Network for Novel View Action Synthesis
    Kara Marie Schatz, Erik Quintanilla, Shruti Vyas, and Yogesh S Rawat
    In Proceedings of the European Conference on Computer Vision (ECCV), 2020
  3. video-action.jpg
    View-invariant action recognition
    Yogesh S Rawat and Shruti Vyas
    In Computer Vision: A Reference Guide, 2020
  4. capsule-net.jpg
    Visual-textual Capsule Routing for Text-based Video Segmentation
    Bruce McIntosh, Kevin Duarte, Yogesh S Rawat, and Mubarak Shah
    In IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2020

2019

  1. capsule-net.jpg
    CapsuleVOS: Semi-Supervised Video Object Segmentation Using Capsule Routing
    Kevin Duarte, Yogesh S Rawat, and Mubarak Shah
    In Proceedings of the IEEE International Conference on Computer Vision, 2019

2018

  1. video-action.jpg
    Time-aware and view-aware video rendering for unsupervised representation learning
    Shruti Vyas, Yogesh S Rawat, and Mubarak Shah
    arXiv preprint arXiv:1811.10699, 2018
  2. capsule-net.jpg
    VideoCapsuleNet: A Simplified Network for Action Detection
    Kevin Duarte, Yogesh S Rawat, and Mubarak Shah
    In Advances in Neural Information Processing Systems, 2018