Collaborative Weakly Supervised Video Correlation Learning for Procedure-Aware Instructional Video Analysis

Tianyao He,Yuxi Li,Huabin Liu,Cheng Zhong,Weiyao Lin,Yang Zhang,Xiao Ma

doi:10.1609/aaai.v38i3.27983

Abstract

Video Correlation Learning (VCL), which aims to analyze the relationships between videos, has been widely studied and applied in various general video tasks. However, applying VCL to instructional videos is still quite challenging due to their intrinsic procedural temporal structure. Specifically, procedural knowledge is critical for accurate correlation analyses on instructional videos. Nevertheless, current procedure-learning methods heavily rely on step-level annotations, which are costly and not scalable. To address this problem, we introduce a weakly supervised framework called Collaborative Procedure Alignment (CPA) for procedure-aware correlation learning on instructional videos. Our framework comprises two core modules: collaborative step mining and frame-to-step alignment. The collaborative step mining module enables simultaneous and consistent step segmentation for paired videos, leveraging the semantic and temporal similarity between frames. Based on the identified steps, the frame-to-step alignment module performs alignment between the frames and steps across videos. The alignment result serves as a measurement of the correlation distance between two videos. We instantiate our framework in two distinct instructional video tasks: sequence verification and action quality assessment. Extensive experiments validate the effectiveness of our approach in providing accurate and interpretable correlation analyses for instructional videos.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Collaborative Weakly Supervised Video Correlation Learning for Procedure-Aware Instructional Video Analysis

Abstract

Talk to us

Similar Papers

More From: Proceedings of the AAAI Conference on Artificial Intelligence

Lead the way for us

Similar Papers

Taking Language out of the Equation: The Assessment of Basic Math Competence Without Language.
Max Greisen ... Claire Muller
Frontiers in Psychology | VOL. 9
Max Greisen, et. al.Max Greisen ... Claire Muller
26 Jun 2018
Frontiers in Psychology | VOL. 9

Order-Constrained Representation Learning for Instructional Video Prediction
Muheng Li ... Jianjiang Feng
IEEE Transactions on Circuits and Systems for Video Technology | VOL. 32
Muheng Li, et. al.Muheng Li ... Jianjiang Feng
01 Aug 2022
IEEE Transactions on Circuits and Systems for Video Technology | VOL. 32

Research on the Impacts of Feedback in Instructional Videos on College Students' Attention and Learning Effects
Xin Lu ... Xue Wang
-
Xin Lu, et. al.Xin Lu ... Xue Wang
05 May 2021
05 May 2021

Immersive virtual reality in orthopedic surgery as elective subject for medical students : First experiences in curricular teaching.
Tobias Schöbel ... Daisy Rotzoll
Orthopadie (Heidelberg, Germany) | VOL. 53
Tobias Schöbel, et. al.Tobias Schöbel ... Daisy Rotzoll
04 Apr 2024
Orthopadie (Heidelberg, Germany) | VOL. 53

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Collaborative Weakly Supervised Video Correlation Learning for Procedure-Aware Instructional Video Analysis

Abstract

Talk to us

Similar Papers

More From: Proceedings of the AAAI Conference on Artificial Intelligence