CASA-Crowd: A Context-Aware Scale Aggregation CNN-Based Crowd Counting Technique

Naveed Ilyas; Ashfaq Ahmad; Kiseon Kim

doi:10.1109/ACCESS.2019.2960292

IEEE Access (Jan 2019)

CASA-Crowd: A Context-Aware Scale Aggregation CNN-Based Crowd Counting Technique

Naveed Ilyas,
Ashfaq Ahmad,
Kiseon Kim

Affiliations

Naveed Ilyas: ORCiD; School of Electrical Engineering and Computer Science, Gwangju Institute of Science and Technology (GIST), Gwangju, South Korea
Ashfaq Ahmad: ORCiD; School of Electrical Engineering and Computing, The University of Newcastle, Callaghan, NSW, Australia
Kiseon Kim: ORCiD; School of Electrical Engineering and Computer Science, Gwangju Institute of Science and Technology (GIST), Gwangju, South Korea

DOI: https://doi.org/10.1109/ACCESS.2019.2960292
Journal volume & issue: Vol. 7
pp. 182050 – 182059

Abstract

Read online

The accuracy of object-based computer vision techniques declines due to major challenges originating from large scale variation, varying shape, perspective variation, and lack of side information. To handle these challenges most of the crowd counting methods use multi-columns (restrict themselves to a set of specific density scenes), deploying a deeper and multi-networks for density estimation. However, these techniques suffer a lot of drawbacks such as extraction of identical features from multi-column, computationally complex architecture, overestimate the density estimation in sparse areas, underestimating in dense areas and averaging of feature maps result in reduced quality of density map. To overcome these drawbacks and to provide a state-of-the-art counting accuracy with comparable computational cost, we therefore propose a deeper and wider network: a Context-aware Scale Aggregation CNN-based Crowd Counting method (CASA-Crowd) to obtain the deep, varying scale and perspective varying features. Further, we include a dilated convolution with varying filter size to obtain contextual information. In addition, due to different dilation rates, a variation in receptive field size is more useful to overcome the perspective distortion. The quality of density map is enhanced while preserving the spatial dimension by obtaining a comparable computational complexity. We further evaluate our method on three well-known datasets: UCF_CC_50, ShanghaiTech Part_A, ShanghaiTech Part_B.

Published in IEEE Access

ISSN: 2169-3536 (Online)
Publisher: IEEE
Country of publisher: United States
LCC subjects: Technology: Electrical engineering. Electronics. Nuclear engineering
Website: https://ieeexplore.ieee.org/xpl/RecentIssue.jsp?punumber=6287639

About the journal

Abstract

Keywords