Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:Unsupervised Action Segmentation with Self-supervised Feature Learning and Co-occurrence Parsing

Jun 02, 2021

Zhe Wang, Hao Chen, Xinyu Li, Chunhui Liu, Yuanjun Xiong, Joseph Tighe, Charless Fowlkes

Figure 1 for Unsupervised Action Segmentation with Self-supervised Feature Learning and Co-occurrence Parsing

Figure 2 for Unsupervised Action Segmentation with Self-supervised Feature Learning and Co-occurrence Parsing

Figure 3 for Unsupervised Action Segmentation with Self-supervised Feature Learning and Co-occurrence Parsing

Figure 4 for Unsupervised Action Segmentation with Self-supervised Feature Learning and Co-occurrence Parsing

Share this with someone who'll enjoy it:

Abstract:Temporal action segmentation is a task to classify each frame in the video with an action label. However, it is quite expensive to annotate every frame in a large corpus of videos to construct a comprehensive supervised training dataset. Thus in this work we explore a self-supervised method that operates on a corpus of unlabeled videos and predicts a likely set of temporal segments across the videos. To do this we leverage self-supervised video classification approaches to perform unsupervised feature extraction. On top of these features we develop CAP, a novel co-occurrence action parsing algorithm that can not only capture the correlation among sub-actions underlying the structure of activities, but also estimate the temporal trajectory of the sub-actions in an accurate and general way. We evaluate on both classic datasets (Breakfast, 50Salads) and emerging fine-grained action datasets (FineGym) with more complex activity structures and similar sub-actions. Results show that our method achieves state-of-the-art performance on all three datasets with up to 22\% improvement, and can even outperform some weakly-supervised approaches, demonstrating its effectiveness and generalizability.

View paper on

Share this with someone who'll enjoy it:

Title:Unsupervised Action Segmentation with Self-supervised Feature Learning and Co-occurrence Parsing

Paper and Code