Papers Dataset Generation
“Dataset Generation” 태그가 달린 논문 308편 · 필터 해제
Vision Language Action Models in Robotic Manipulation: A Systematic Review
Vision Language Action (VLA) models represent a transformative shift in robotics, with the aim of unifying visual perception, natural language understanding, and embodied control within a single learning framework. This …
Dataset GenerationNatural Language UnderstandingVision-Language-ActionCommunicating Smartly in the Molecular Domain: Neural Networks in the Internet of Bio-Nano Things
Recent developments in the Internet of Bio-Nano Things (IoBNT) are laying the groundwork for innovative applications across the healthcare sector. Nanodevices designed to operate within the body, managed remotely via the…
Dataset GenerationExplainable artificial intelligenceExplainable Artificial Intelligence (XAI)From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation
Current research on long-form context in Large Language Models (LLMs) primarily focuses on the understanding of long-contexts, the Open-ended Long Text Generation (Open-LTG) remains insufficiently explored. Training a lo…
Dataset GenerationReinforcement Learning (RL)Text GenerationEnhancing Clinical Models with Pseudo Data for De-identification
Many models are pretrained on redacted text for privacy reasons. Clinical foundation models are often trained on de-identified text, which uses special syntax (masked) text in place of protected health information. Even …
Dataset GenerationDe-identificationA large-scale, physically-based synthetic dataset for satellite pose estimation
The Deep Learning Visual Space Simulation System (DLVS3) introduces a novel synthetic dataset generator and a simulation pipeline specifically designed for training and testing satellite pose estimation solutions. This w…
BenchmarkingDataset GenerationDeep LearningPose Estimation+1Real-Time Per-Garment Virtual Try-On with Temporal Consistency for Loose-Fitting Garments
Per-garment virtual try-on methods collect garment-specific datasets and train networks tailored to each garment to achieve superior results. However, these approaches often struggle with loose-fitting garments due to tw…
Dataset GenerationVirtual Try-onCode Execution as Grounded Supervision for LLM Reasoning
Training large language models (LLMs) with chain-of-thought (CoT) supervision has proven effective for enhancing their reasoning abilities. However, obtaining reliable and accurate reasoning supervision remains a signifi…
Dataset GenerationHierarchical Lexical Graph for Enhanced Multi-Hop Retrieval
Retrieval-Augmented Generation (RAG) grounds large language models in external evidence, yet it still falters when answers must be pieced together across semantically distant documents. We close this gap with the Hierarc…
Dataset GenerationRAGRetrievalRetrieval-augmented GenerationSynthetic Dataset Generation for Autonomous Mobile Robots Using 3D Gaussian Splatting for Vision Training
Annotated datasets are critical for training neural networks for object detection, yet their manual creation is time- and labour-intensive, subjective to human error, and often limited in diversity. This challenge is par…
Dataset Generationobject-detectionObject DetectionSynthetic Data GenerationGenerating Synthetic Stereo Datasets using 3D Gaussian Splatting and Expert Knowledge Transfer
In this paper, we introduce a 3D Gaussian Splatting (3DGS)-based pipeline for stereo dataset generation, offering an efficient alternative to Neural Radiance Fields (NeRF)-based methods. To obtain useful geometry estimat…
3DGSDataset GenerationNeRFTransfer Learning+1MORSE-500: A Programmatically Controllable Video Benchmark to Stress-Test Multimodal Reasoning
Despite rapid advances in vision-language models (VLMs), current benchmarks for multimodal reasoning fall short in three key dimensions. First, they overwhelmingly rely on static images, failing to capture the temporal c…
Dataset GenerationMathematical Problem-SolvingMultimodal ReasoningCETBench: A Novel Dataset constructed via Transformations over Programs for Benchmarking LLMs for Code-Equivalence Checking
LLMs have been extensively used for the task of automated code generation. In this work, we examine the applicability of LLMs for the related but relatively unexplored task of code-equivalence checking, i.e., given two p…
BenchmarkingCode GenerationCode TranslationDataset GenerationTimeGraph: Synthetic Benchmark Datasets for Robust Time-Series Causal Discovery
Robust causal discovery in time series datasets depends on reliable benchmark datasets with known ground-truth causal relationships. However, such datasets remain scarce, and existing synthetic alternatives often overloo…
Causal DiscoveryDataset GenerationTime SeriesMulti-Domain ABSA Conversation Dataset Generation via LLMs for Real-World Evaluation and Model Comparison
Aspect-Based Sentiment Analysis (ABSA) offers granular insights into opinions but often suffers from the scarcity of diverse, labeled datasets that reflect real-world conversational nuances. This paper presents an approa…
Aspect-Based Sentiment AnalysisAspect-Based Sentiment Analysis (ABSA)Dataset GenerationSentiment Analysis+2Track Anything Annotate: Video annotation and dataset generation of computer vision models
Modern machine learning methods require significant amounts of labelled data, making the preparation process time-consuming and resource-intensive. In this paper, we propose to consider the process of prototyping a tool …
Dataset GenerationF-ANcGAN: An Attention-Enhanced Cycle Consistent Generative Adversarial Architecture for Synthetic Image Generation of Nanoparticles
Nanomaterial research is becoming a vital area for energy, medicine, and materials science, and accurate analysis of the nanoparticle topology is essential to determine their properties. Unfortunately, the lack of high-q…
Dataset GenerationImage GenerationSegmentationRobo2VLM: Visual Question Answering from Large-Scale In-the-Wild Robot Manipulation Datasets
Vision-Language Models (VLMs) acquire real-world knowledge and general reasoning ability through Internet-scale image-text corpora. They can augment robotic systems with scene understanding and task planning, and assist …
Dataset GenerationDescriptiveMultiple-choiceQuestion Answering+6seg_3D_by_PC2D: Multi-View Projection for Domain Generalization and Adaptation in 3D Semantic Segmentation
3D semantic segmentation plays a pivotal role in autonomous driving and road infrastructure analysis, yet state-of-the-art 3D models are prone to severe domain shift when deployed across different datasets. We propose a …
3D Semantic SegmentationAutonomous DrivingDataset GenerationDomain Adaptation+3FMSD-TTS: Few-shot Multi-Speaker Multi-Dialect Text-to-Speech Synthesis for Ü-Tsang, Amdo and Kham Speech Dataset Generation
Tibetan is a low-resource language with minimal parallel speech corpora spanning its three major dialects-\"U-Tsang, Amdo, and Kham-limiting progress in speech modeling. To address this issue, we propose FMSD-TTS, a few-…
Dataset GenerationSpeech Synthesistext-to-speechText to Speech+1Automatic Dataset Generation for Knowledge Intensive Question Answering Tasks
A question-answering (QA) system is to search suitable answers within a knowledge base. Current QA systems struggle with queries requiring complex reasoning or real-time knowledge integration. They are often supplemented…
Dataset GenerationQuestion AnsweringRAGRetrieval+1