• Complex
  • Title
  • Keyword
  • Abstract
  • Scholars
  • Journal
  • ISSN
  • Conference
搜索

Author:

Zhang, Chengyang (Zhang, Chengyang.) | Zhang, Yong (Zhang, Yong.) | Li, Bo (Li, Bo.) | Piao, Xinglin (Piao, Xinglin.) | Yin, Baocai (Yin, Baocai.)

Indexed by:

EI Scopus SCIE

Abstract:

Most existing weakly supervised crowd counting methods utilize Convolutional Neural Networks (CNN) or Transformer to estimate the total number of individuals in an image. However, both CNN-based (grid-to-count paradigm) and Transformer-based (sequence-to-count paradigm) methods take images as inputs in a regular form. This approach treats all pixels equally but cannot address the uneven distribution problem within human crowds. This challenge would lead to a decline in the counting performance of the model. Compared with grid and sequence, the graph structure could better explore the relationship among features. In this article, we propose a new graph-based crowd counting method named CrowdGraph, which reinterprets the weakly supervised crowd counting problem from a graph-to-count perspective. In the proposed CrowdGraph, each image is constructed as a graph, and a graph-based network is designed to extract features at the graph level. CrowdGraph comprises three main components: a dynamic graph convolutional backbone, a multi-scale dilated graph convolution module, and a regression head. To the best of our knowledge, CrowdGraph is the first method that is completely formulated based on the Graph Neural Network (GNN) for the crowd counting task. Extensive experiments demonstrate that the proposed CrowdGraph outperforms pure CNN-based and pure Transformer-based weakly supervised methods comprehensively and achieves highly competitive counting performance. © 2024 Copyright held by the owner/author(s). Publication rights licensed to ACM.

Keyword:

Graph neural networks Convolution Supervised learning Convolutional neural networks Graphic methods

Author Community:

  • [ 1 ] [Zhang, Chengyang]Beijing Key Laboratory of Multimedia and Intelligent Software Technology, Beijing Artiicial Intelligence Institute, Faculty of Information Technology, Beijing University of Technology, Beijing; 100124, China
  • [ 2 ] [Zhang, Yong]Beijing Key Laboratory of Multimedia and Intelligent Software Technology, Beijing Artiicial Intelligence Institute, Faculty of Information Technology, Beijing University of Technology, Beijing; 100124, China
  • [ 3 ] [Li, Bo]Beijing Key Laboratory of Multimedia and Intelligent Software Technology, Beijing Artiicial Intelligence Institute, Faculty of Information Technology, Beijing University of Technology, Beijing; 100124, China
  • [ 4 ] [Piao, Xinglin]Beijing Key Laboratory of Multimedia and Intelligent Software Technology, Beijing Artiicial Intelligence Institute, Faculty of Information Technology, Beijing University of Technology, Beijing; 100124, China
  • [ 5 ] [Yin, Baocai]Beijing Key Laboratory of Multimedia and Intelligent Software Technology, Beijing Artiicial Intelligence Institute, Faculty of Information Technology, Beijing University of Technology, Beijing; 100124, China

Reprint Author's Address:

Email:

Show more details

Related Keywords:

Related Article:

Source :

ACM Transactions on Multimedia Computing, Communications and Applications

ISSN: 1551-6857

Year: 2024

Issue: 5

Volume: 20

5 . 1 0 0

JCR@2022

Cited Count:

WoS CC Cited Count: 0

SCOPUS Cited Count: 5

ESI Highly Cited Papers on the List: 0 Unfold All

WanFang Cited Count:

Chinese Cited Count:

30 Days PV: 7

Affiliated Colleges:

Online/Total:603/10599306
Address:BJUT Library(100 Pingleyuan,Chaoyang District,Beijing 100124, China Post Code:100124) Contact Us:010-67392185
Copyright:BJUT Library Technical Support:Beijing Aegean Software Co., Ltd.