paper-with-me

홈 › Papers

Towards Data-Centric AI: A Comprehensive Survey of Traditional, Reinforcement, and Generative Approaches for Tabular Data Transformation

2025-01-17 · Dongjie Wang, Yanyong Huang, Wangyang Ying, Haoyue Bai, Nanxu Gong, Xinyuan Wang, Sixun Dong, Tao Zhe, Kunpeng Liu, Meng Xiao, Pengfei Wang, Pengyang Wang, Hui Xiong, Yanjie Fu

Tabular data is one of the most widely used formats across industries, driving critical applications in areas such as finance, healthcare, and marketing. In the era of data-centric AI, improving data quality and representation has become essential for enhancing model performance, particularly in applications centered around tabular data. This survey examines the key aspects of tabular data-centric AI, emphasizing feature selection and feature generation as essential techniques for data space refinement. We provide a systematic review of feature selection methods, which identify and retain the most relevant data attributes, and feature generation approaches, which create new features to simplify the capture of complex data patterns. This survey offers a comprehensive overview of current methodologies through an analysis of recent advancements, practical applications, and the strengths and limitations of these techniques. Finally, we outline open challenges and suggest future perspectives to inspire continued innovation in this field.

📄 PDF Abstract BibTeX arXiv:2501.10555

Code (0)

등록된 구현이 없습니다.

Tasks

feature selectionMarketingSurvey

Methods 이 논문이 사용한 방법론

Feature Selection Feature selection, also known as variable selection, attribute selection or variable subset selection, is the process of selecting a subset of relevant features (variables,…

Similar Papers 제목 키워드 기반

A Survey of Reinforcement Learning for Large Language Models under Data Scarcity: Challenges and Solutions

2026-04-19 · Zhiyin Yu, Yuchen Mou, Juncheng Yan, Junyu Luo 외 arxiv

Reinforcement learning (RL) has emerged as a powerful post-training paradigm for enhancing the reasoning capabilities of large language models (LLMs). However, reinforcement learning for LLMs faces substantial data scarc…

Reinforcement Learning

Sensing, Social, and Motion Intelligence in Embodied Navigation: A Comprehensive Survey

2025-08-21 · Chaoran Xiong, Yulong Huang, Fangwen Yu, Changhao Chen 외 arxiv

Embodied navigation (EN) advances traditional navigation by enabling robots to perform complex egocentric tasks through sensing, social, and motion intelligence. In contrast to classic methodologies that rely on explicit…

Closer Look at Efficient Inference Methods: A Survey of Speculative Decoding

2024-11-20 · Hyun Ryu, Eric Kim

Efficient inference in large language models (LLMs) has become a critical focus as their scale and complexity grow. Traditional autoregressive decoding, while effective, suffers from computational inefficiencies due to i…

Survey

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling

2025-02-12 · Shixiang Tang, Yizhou Wang, Lu Chen, YuAn Wang 외

Human understanding and generation are critical for modeling digital humans and humanoid embodiments. Recently, Human-centric Foundation Models (HcFMs) inspired by the success of generalist models, such as large language…

Survey

A Survey on Human-Centric LLMs

2024-11-20 · Jing Yi Wang, Nicholas Sukiennik, Tong Li, Weikang Su 외

The rapid evolution of large language models (LLMs) and their capacity to simulate human cognition and behavior has given rise to LLM-based frameworks and tools that are evaluated and applied based on their ability to pe…

Decision MakingEmotional IntelligenceSociologySurvey