Deep Reinforcement Learning for Query-Conditioned Video Summarization

Yujia Zhang; Michael Kampffmeyer; Xiaoguang Zhao; Min Tan

doi:10.3390/app9040750

Applied Sciences (Feb 2019)

Deep Reinforcement Learning for Query-Conditioned Video Summarization

Yujia Zhang,
Michael Kampffmeyer,
Xiaoguang Zhao,
Min Tan

Affiliations

Yujia Zhang: Institute of Automation, Chinese Academy of Sciences, Beijing 100190, China
Michael Kampffmeyer: Machine Learning Group, UiT The Arctic University of Norway, Tromsø 9019, Norway
Xiaoguang Zhao: Institute of Automation, Chinese Academy of Sciences, Beijing 100190, China
Min Tan: Institute of Automation, Chinese Academy of Sciences, Beijing 100190, China

DOI: https://doi.org/10.3390/app9040750
Journal volume & issue: Vol. 9, no. 4
p. 750

Abstract

Read online

Query-conditioned video summarization requires to (1) find a diverse set of video shots/frames that are representative for the whole video, and that (2) the selected shots/frames are related to a given query. Thus it can be tailored to different user interests leading to a better personalized summary and differs from the generic video summarization which only focuses on video content. Our work targets this query-conditioned video summarization task, by first proposing a Mapping Network (MapNet) in order to express how related a shot is to a given query. MapNet helps establish the relation between the two different modalities (videos and query), which allows mapping of visual information to query space. After that, a deep reinforcement learning-based summarization network (SummNet) is developed to provide personalized summaries by integrating relatedness, representativeness and diversity rewards. These rewards jointly guide the agent to select the most representative and diversity video shots that are most related to the user query. Experimental results on a query-conditioned video summarization benchmark demonstrate the effectiveness of our proposed method, indicating the usefulness of the proposed mapping mechanism as well as the reinforcement learning approach.

Published in Applied Sciences

ISSN: 2076-3417 (Online)
Publisher: MDPI AG
Country of publisher: Switzerland
LCC subjects: Technology: Engineering (General). Civil engineering (General); Science: Biology (General); Science: Physics; Science: Chemistry
Website: http://www.mdpi.com/journal/applsci

About the journal

Abstract

Keywords