跳到主要导航 跳到搜索 跳到主要内容

Semantic Preprocessor for Image Compression for Machines

  • Mingyi Yang
  • , Luis Herranz
  • , Fei Yang
  • , Luka Murn
  • , Marc Gorriz Blanch
  • , Shuai Wan
  • , Fuzheng Yang
  • , Marta Mrak
  • Xidian University
  • Computer Vision Centre
  • Autonomous University of Barcelona
  • BBC

科研成果: 期刊稿件会议文章同行评审

3 引用 (Scopus)

摘要

Visual content is being increasingly transmitted and consumed by machines rather than humans to perform automated content analysis tasks. In this paper, we propose an image preprocessor that optimizes the input image for machine consumption prior to encoding by an off-the-shelf codec designed for human consumption. To achieve a better trade-off between the accuracy of the machine analysis task and bitrate, we propose leveraging pre-extracted semantic information to improve the preprocessor's ability to accurately identify and filter out task-irrelevant information. Furthermore, we propose a two-part loss function to optimize the preprocessor, consisted of a rate-task performance loss and a semantic distillation loss, which helps the reconstructed image obtain more information that contributes to the accuracy of the task. Experiments show that the proposed preprocessor can save up to 48.83% bitrate compared with the method without the preprocessor, and save up to 36.24% bitrate compared to existing preprocessors for machine vision.

学术指纹

探究 'Semantic Preprocessor for Image Compression for Machines' 的科研主题。它们共同构成独一无二的学术指纹。

引用此