Automatic Generation of Two Level Hierarchical Tutorials from Instructional Makeup Videos

We present a multi-modal approach for automatically generating hierarchical tutorials from instructional makeup videos. Our approach is inspired by prior research in cognitive psychology, which suggests that people mentally segment procedural tasks into event hierarchies, where coarse-grained events focus on objects while fine-grained events focus on actions. In the instructional makeup domain, we find that objects correspond to facial parts while fine-grained steps correspond to actions on those facial parts. Given an input instructional makeup video, we apply a set of heuristics that combine computer vision techniques with transcript text analysis to automatically identify the fine-level action steps and group these steps by facial part to form the coarse-level events. We provide a voice-enabled, mixed-media UI to visualize the resulting hierarchy and allow users to efficiently navigate the tutorial (e.g., skip ahead, return to previous steps) at their own pace. Users can navigate the hierarchy at both the facial-part and action-step levels using click-based interactions and voice commands. We demonstrate the effectiveness of segmentation algorithms and the resulting mixed-media UI on a variety of input makeup videos. A user study shows that users prefer following instructional makeup videos in our mixed-media format to the standard video UI and that they find our format much easier to navigate.

Stanford University, Stanford, California, United States

Google Research, Mountain View, California, United States

Google Research, San Francisco, California, United States

Google, Atlanta, Georgia, United States

Stanford University, Stanford, California, United States

10.1145/3411764.3445721

https://doi.org/10.1145/3411764.3445721

The ACM CHI Conference on Human Factors in Computing Systems (https://chi2021.acm.org/)

[B] Paper Room 05, 2021-05-14 01:00:00~2021-05-14 03:00:00 / [C] Paper Room 05, 2021-05-14 09:00:00~2021-05-14 11:00:00 / [A] Paper Room 05, 2021-05-13 17:00:00~2021-05-13 19:00:00

Paper Room 05

14 件の発表

開始日時2021-05-14 01:00:00

終了日時2021-05-14 03:00:00

読み込み中…

お気に入り

あとで読む

コレクション

Automatic Generation of Two Level Hierarchical Tutorials from Instructional Makeup Videos

要旨

著者

DOI

論文URL

動画

会議: CHI 2021

セッション: Engineering Interactive Applications

日本語まとめ