paper-with-me

홈 › Papers

A Style-Based Profiling Framework for Quantifying the Synthetic-to-Real Gap in Autonomous Driving Datasets

2025-10-11 · Dingyi Yao, Xinyao Han, Ruibo Ming, Zhihang Song, Lihui Peng, Jianming Hu, Danya Yao, Yi Zhang arxiv

Ensuring the reliability of autonomous driving perception systems requires extensive environment-based testing, yet real-world execution is often impractical. Synthetic datasets have therefore emerged as a promising alternative, offering advantages such as cost-effectiveness, bias free labeling, and controllable scenarios. However, the domain gap between synthetic and real-world datasets remains a major obstacle to model generalization. To address this challenge from a data-centric perspective, this paper introduces a profile extraction and discovery framework for characterizing the style profiles underlying both synthetic and real image datasets. We propose Style Embedding Distribution Discrepancy (SEDD) as a novel evaluation metric. Our framework combines Gram matrix-based style extraction with metric learning optimized for intra-class compactness and inter-class separation to extract style embeddings. Furthermore, we establish a benchmark using publicly available datasets. Experiments are conducted on a variety of datasets and sim-to-real methods, and the results show that our method is capable of quantifying the synthetic-to-real gap. This work provides a standardized profiling-based quality control paradigm that enables systematic diagnosis and targeted enhancement of synthetic datasets, advancing future development of data-driven autonomous driving systems.

📄 PDF Abstract BibTeX arXiv:2510.10203

Code (0)

등록된 구현이 없습니다.

Tasks

Autonomous DrivingMetric Learning

Similar Papers 제목 키워드 기반

HumAIne-Chatbot: Real-Time Personalized Conversational AI via Reinforcement Learning

2025-09-04 · Georgios Makridis, George Fragiadakis, Jorge Oliveira, Tomaz Saraiva 외 arxiv

Current conversational AI systems often provide generic, one-size-fits-all interactions that overlook individual user characteristics and lack adaptive dialogue management. To address this gap, we introduce \textbf{HumAI…

Reinforcement Learning

ProfiLLM: An LLM-Based Framework for Implicit Profiling of Chatbot Users

2025-06-16 · Shahaf David, Yair Meidan, Ido Hersko, Daniel Varnovitzky 외

Despite significant advancements in conversational AI, large language model (LLM)-powered chatbots often struggle with personalizing their responses according to individual user characteristics, such as technical experti…

ChatbotLarge Language Model

LLM-Generated Negative News Headlines Dataset: Creation and Benchmarking Against Real Journalism

2025-10-24 · Olusola Babalola, Bolanle Ojokoh, Olutayo Boyinbode arxiv

This research examines the potential of datasets generated by Large Language Models (LLMs) to support Natural Language Processing (NLP) tasks, aiming to overcome challenges related to data acquisition and privacy concern…

Semantic SimilaritySentiment Analysis

A framework for mining lifestyle profiles through multi-dimensional and high-order mobility feature clustering

2023-12-01 · Yeshuo Shu, Gangcheng Zhang, Keyi Liu, Jintong Tang 외

Human mobility demonstrates a high degree of regularity, which facilitates the discovery of lifestyle profiles. Existing research has yet to fully utilize the regularities embedded in high-order features extracted from h…

Common Sense ReasoningFeature EngineeringTime Series

RCScore: Quantifying Response Consistency in Large Language Models

2025-10-30 · Dongjun Jang, Youngchae Ahn, Hyopil Shin arxiv

Current LLM evaluations often rely on a single instruction template, overlooking models' sensitivity to instruction style-a critical aspect for real-world deployments. We present RCScore, a multi-dimensional framework qu…