摘要
Video has become one of the most important data forms. Video distillation explores more compact data forms and information modalities by analyzing the spatial-temporal and semantic features of video data, which is an important task in computer vision and a key technique in artificial intelligence. With the rapid development of video capturing devices and the increasing human requirements, video analysis tasks are facing numbers of opportunities and challenges. In recent years, large amounts of video distillation approaches are proposed. This paper creatively unifies the theoretical basis of video distillation by analyzing the relationship among data, information and knowledge from the perspective of information theory, and argues that the principle of video distillation is to improve the information capacity of video data. Then, we overview existing approaches in the aspects of video data representation, key content summarization, moving object synopsis and text description generation, etc., and relate the development of video summarization, synopsis and captioning, which are typical tasks in video distillation. More importantly, this paper discusses the advantages and drawbacks of existing approaches, and then points out several key scientific problems that have not yet been addressed, and simultaneously analyzes the potential future development in video distillation.
| 投稿的翻译标题 | Video distillation |
|---|---|
| 源语言 | 繁体中文 |
| 页(从-至) | 695-734 |
| 页数 | 40 |
| 期刊 | Scientia Sinica Informationis |
| 卷 | 51 |
| 期 | 5 |
| DOI | |
| 出版状态 | 已出版 - 5月 2021 |
关键词
- Artificial intelligence
- Computer vision
- Video captioning
- Video distillation
- Video summarization
- Video synopsis
- Visual representation
指纹
探究 '视频萃取' 的科研主题。它们共同构成独一无二的指纹。引用此
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver