paper-with-me

홈 › Papers

SynPlay: Importing Real-world Diversity for a Synthetic Human Dataset

2024-08-21 · Jinsub Yim, Hyungtae Lee, Sungmin Eum, Yi-Ting Shen, Yan Zhang, Heesung Kwon, Shuvra S. Bhattacharyya

We introduce Synthetic Playground (SynPlay), a new synthetic human dataset that aims to bring out the diversity of human appearance in the real world. We focus on two factors to achieve a level of diversity that has not yet been seen in previous works: i) realistic human motions and poses and ii) multiple camera viewpoints towards human instances. We first use a game engine and its library-provided elementary motions to create games where virtual players can take less-constrained and natural movements while following the game rules (i.e., rule-guided motion design as opposed to detail-guided design). We then augment the elementary motions with real human motions captured with a motion capture device. To render various human appearances in the games from multiple viewpoints, we use seven virtual cameras encompassing the ground and aerial views, capturing abundant aerial-vs-ground and dynamic-vs-static attributes of the scene. Through extensive and carefully-designed experiments, we show that using SynPlay in model training leads to enhanced accuracy over existing synthetic datasets for human detection and segmentation. The benefit of SynPlay becomes even greater for tasks in the data-scarce regime, such as few-shot and cross-domain learning tasks. These results clearly demonstrate that SynPlay can be used as an essential dataset with rich attributes of complex human appearances and poses suitable for model pretraining. SynPlay dataset comprising over 73k images and 6.5M human instances, is available for download at https://synplaydataset.github.io/.

📄 PDF Abstract BibTeX arXiv:2408.11814

Code (0)

등록된 구현이 없습니다.

Tasks

DiversityHuman Detection

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

ModTrans: Translating Real-world Models for Distributed Training Simulator

2026-04-02 · Yi Lyu arxiv

Large-scale distributed training has been a research hot spot in machine learning systems for industry and academia in recent years. However, conducting experiments without physical machines and corresponding resources i…

Exploring Precision and Recall to assess the quality and diversity of LLMs

2024-02-16 · Florian Le Bronnec, Alexandre Verine, Benjamin Negrevergne, Yann Chevaleyre 외

We introduce a novel evaluation framework for Large Language Models (LLMs) such as \textsc{Llama-2} and \textsc{Mistral}, focusing on importing Precision and Recall metrics from image generation to text generation. This …

DiversityImage GenerationText Generation

Wind Turbine Feature Detection Using Deep Learning and Synthetic Data

2025-07-29 · Arash Shahirpour, Jakob Gebler, Manuel Sanders, Tim Reuscher arxiv

For the autonomous drone-based inspection of wind turbine (WT) blades, accurate detection of the WT and its key features is essential for safe drone positioning and collision avoidance. Existing deep learning methods typ…

Collision Avoidance

Bridging the Gap: Enhancing the Utility of Synthetic Data via Post-Processing Techniques

2023-05-17 · Andrea Lampis, Eugenio Lomurno, Matteo Matteucci

Acquiring and annotating suitable datasets for training deep learning models is challenging. This often results in tedious and time-consuming efforts that can hinder research progress. However, generative models have eme…

Diversity

Learning to Import through Production Networks

2024-05-22 · Kenan Huremović, Federico Nutarelli, Francesco Serti, Fernando Vega-Redondo

Using administrative data on the universe of inter-firm transactions in Spain, we show that firms learn to import from their domestic suppliers and customers. Our identification strategy exploits the panel structure of t…