We introduce the novel task of Crowd Volume Estimation (CVE), defined as the process of estimating the collective body volume of crowds using only RGB images. Besides event management and public safety, CVE can be instrumental in approximating body weight, unlocking weight-sensitive applications such as infrastructure stress assessment, and assuring even weight balance. We propose the first benchmark for CVE, comprising ANTHROPOS-V, a synthetic photorealistic video dataset featuring crowds in diverse urban environments. Its annotations include each person's volume, SMPL shape parameters, and key-points. Also, we explore metrics pertinent to CVE, define baseline models adapted from Human Mesh Recovery and Crowd Counting domains, and propose a CVE-specific methodology that surpasses baselines. Although synthetic, the weights and heights of individuals are aligned with the real-world population distribution across gen-ders, and they transfer to the downstream task of CVE from real images. Benchmark and code are available at github.com/colloronelucaICrowd-Volume-Estimation.

ANTHROPOS-V: Benchmarking the Novel Task of Crowd Volume Estimation / Collorone, Luca; D'Arrigo, Stefano; Pappa, Massimiliano; D'Amely Di Melendugno, Guido M.; Ficarra, Giovanni; Galasso, Fabio. - (2025), pp. 5284-5294. (Intervento presentato al convegno 2025 IEEE/CVF Winter Conference on Applications of Computer Vision, WACV 2025 tenutosi a Tucson; Usa (AZ)) [10.1109/wacv61041.2025.00516].

ANTHROPOS-V: Benchmarking the Novel Task of Crowd Volume Estimation

Collorone, Luca
;
D'Arrigo, Stefano;Pappa, Massimiliano;D'Amely Di Melendugno, Guido M.;Ficarra, Giovanni;Galasso, Fabio
2025

Abstract

We introduce the novel task of Crowd Volume Estimation (CVE), defined as the process of estimating the collective body volume of crowds using only RGB images. Besides event management and public safety, CVE can be instrumental in approximating body weight, unlocking weight-sensitive applications such as infrastructure stress assessment, and assuring even weight balance. We propose the first benchmark for CVE, comprising ANTHROPOS-V, a synthetic photorealistic video dataset featuring crowds in diverse urban environments. Its annotations include each person's volume, SMPL shape parameters, and key-points. Also, we explore metrics pertinent to CVE, define baseline models adapted from Human Mesh Recovery and Crowd Counting domains, and propose a CVE-specific methodology that surpasses baselines. Although synthetic, the weights and heights of individuals are aligned with the real-world population distribution across gen-ders, and they transfer to the downstream task of CVE from real images. Benchmark and code are available at github.com/colloronelucaICrowd-Volume-Estimation.
2025
2025 IEEE/CVF Winter Conference on Applications of Computer Vision, WACV 2025
crowd counting; crowd volume estimation; human anthropometrics; human mesh recovery
04 Pubblicazione in atti di convegno::04b Atto di convegno in volume
ANTHROPOS-V: Benchmarking the Novel Task of Crowd Volume Estimation / Collorone, Luca; D'Arrigo, Stefano; Pappa, Massimiliano; D'Amely Di Melendugno, Guido M.; Ficarra, Giovanni; Galasso, Fabio. - (2025), pp. 5284-5294. (Intervento presentato al convegno 2025 IEEE/CVF Winter Conference on Applications of Computer Vision, WACV 2025 tenutosi a Tucson; Usa (AZ)) [10.1109/wacv61041.2025.00516].
File allegati a questo prodotto
Non ci sono file associati a questo prodotto.

I documenti in IRIS sono protetti da copyright e tutti i diritti sono riservati, salvo diversa indicazione.

Utilizza questo identificativo per citare o creare un link a questo documento: https://hdl.handle.net/11573/1741969
 Attenzione

Attenzione! I dati visualizzati non sono stati sottoposti a validazione da parte dell'ateneo

Citazioni
  • ???jsp.display-item.citation.pmc??? ND
  • Scopus 0
  • ???jsp.display-item.citation.isi??? ND
social impact