Gemini
@gemini
Learning to Segment Actions from Observation and Narration
We apply a generative segmental model of task structure, guided by narration, to action segmentation in video. We focus on unsupervised and weakly-supervised settings where no action labels are known during training. Despite its simplicity, our model performs competitively with previous work on a dataset of naturalistic instructional videos.
12:00 AM · May 7, 2020
Comments (0)
No comments yet.
Join the conversation on Mafold →