Closing the Gap in Human Behavior Analysis: A Pipeline for Synthesizing Trimodal Data

Stippel, Christian; Heitzinger, Thomas; Sterzinger, Rafael; Kampel, Martin

doi:10.1109/PerComWorkshops59983.2024.10503351

Computer Science > Computer Vision and Pattern Recognition

arXiv:2402.01537 (cs)

[Submitted on 2 Feb 2024]

Title:Closing the Gap in Human Behavior Analysis: A Pipeline for Synthesizing Trimodal Data

Authors:Christian Stippel, Thomas Heitzinger, Rafael Sterzinger, Martin Kampel

View PDF

Abstract:In pervasive machine learning, especially in Human Behavior Analysis (HBA), RGB has been the primary modality due to its accessibility and richness of information. However, linked with its benefits are challenges, including sensitivity to lighting conditions and privacy concerns. One possibility to overcome these vulnerabilities is to resort to different modalities. For instance, thermal is particularly adept at accentuating human forms, while depth adds crucial contextual layers. Despite their known benefits, only a few HBA-specific datasets that integrate these modalities exist. To address this shortage, our research introduces a novel generative technique for creating trimodal, i.e., RGB, thermal, and depth, human-focused datasets. This technique capitalizes on human segmentation masks derived from RGB images, combined with thermal and depth backgrounds that are sourced automatically. With these two ingredients, we synthesize depth and thermal counterparts from existing RGB data utilizing conditional image-to-image translation. By employing this approach, we generate trimodal data that can be leveraged to train models for settings with limited data, bad lightning conditions, or privacy-sensitive areas.

Subjects:	Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
Cite as:	arXiv:2402.01537 [cs.CV]
	(or arXiv:2402.01537v1 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2402.01537
Related DOI:	https://doi.org/10.1109/PerComWorkshops59983.2024.10503351

Submission history

From: Rafael Sterzinger [view email]
[v1] Fri, 2 Feb 2024 16:27:45 UTC (3,683 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Closing the Gap in Human Behavior Analysis: A Pipeline for Synthesizing Trimodal Data

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Closing the Gap in Human Behavior Analysis: A Pipeline for Synthesizing Trimodal Data

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators