跳到主要导航 跳到搜索 跳到主要内容

Body structure aware deep crowd counting

  • Siyu Huang
  • , Xi Li
  • , Zhongfei Zhang
  • , Fei Wu
  • , Shenghua Gao
  • , Rongrong Ji
  • , Junwei Han
  • Zhejiang University
  • State University of New York Binghamton University
  • ShanghaiTech University
  • Xiamen University

科研成果: 期刊稿件文章同行评审

113 引用 (Scopus)

摘要

Crowd counting is a challenging task, mainly due to the severe occlusions among dense crowds. This paper aims to take a broader view to address crowd counting from the perspective of semantic modeling. In essence, crowd counting is a task of pedestrian semantic analysis involving three key factors: pedestrians, heads, and their context structure. The information of different body parts is an important cue to help us judge whether there exists a person at a certain position. Existing methods usually perform crowd counting from the perspective of directly modeling the visual properties of either the whole body or the heads only, without explicitly capturing the composite body-part semantic structure information that is crucial for crowd counting. In our approach, we first formulate the key factors of crowd counting as semantic scene models. Then, we convert the crowd counting problem into a multi-task learning problem, such that the semantic scene models are turned into different sub-tasks. Finally, the deep convolutional neural networks are used to learn the sub-tasks in a unified scheme. Our approach encodes the semantic nature of crowd counting and provides a novel solution in terms of pedestrian semantic analysis. In experiments, our approach outperforms the state-ofthe- art methods on four benchmark crowd counting data sets. The semantic structure information is demonstrated to be an effective cue in scene of crowd counting.

源语言英语
页(从-至)1049-1059
页数11
期刊IEEE Transactions on Image Processing
27
3
DOI
出版状态已出版 - 3月 2018

指纹

探究 'Body structure aware deep crowd counting' 的科研主题。它们共同构成独一无二的指纹。

引用此