Temporal Localization of Fine-Grained Actions in Videos by Domain Transfer from Web Images

10 years 13 days ago

Download www.cs.cmu.edu

We address the problem of ﬁne-grained action localization from temporally untrimmed web videos. We assume that only weak video-level annotations are available for training. The goal is to use these weak labels to identify temporal segments corresponding to the actions, and learn models that generalize to unconstrained web videos. We ﬁnd that web images queried by action names serve as well-localized highlights for many actions, but are noisily labeled. To solve this problem, we propose a simple yet eﬀective method that takes weak video labels and noisy image labels as input, and generates localized action frames as output. This is achieved by cross-domain transfer between video frames and web images, using pre-trained deep convolutional neural networks. We then use the localized action frames to train action recognition models with long short-term memory networks. We collect a ﬁne-grained sports action data set FGA-240 of more than 130,000 YouTube videos. It has 240 ﬁne-grai...

Chen Sun, Sanketh Shetty, Rahul Sukthankar, Ram Ne

Real-time Traffic

MM 2015 | Multimedia |

claim paper

Added	14 Apr 2016
Updated	14 Apr 2016
Type	Journal
Year	2015
Where	MM
Authors	Chen Sun, Sanketh Shetty, Rahul Sukthankar, Ram Nevatia

Sciweavers

Temporal Localization of Fine-Grained Actions in Videos by Domain Transfer from Web Images

MM 2015 | Multimedia |

Explore & Download

Productivity Tools

Sciweavers