Are Large Language Models the New Interface for Data Pipelines?
A Language Model is a term that encompasses various types of models designed to understand and generate human communication. Large Language Models (LLMs) have gained significant attention due to their ability to process text with human-like fluency and coherence, making them valuable for a wide range of data-related tasks fashioned as pipelines. The capabilities of LLMs in natural language understanding and generation, combined with their scalability, versatility, and state-of-the-art performance, enable innovative applications across various AI-related fields, including eXplainable Artificial Intelligence (XAI), Automated Machine Learning (AutoML), and Knowledge Graphs (KG). Furthermore, we believe these models can extract valuable insights and make data-driven decisions at scale, a practice commonly referred to as Big Data Analytics (BDA). In this position paper, we provide some discussions in the direction of unlocking synergies among these technologies, which can lead to more powerful and intelligent AI solutions, driving improvements in data pipelines across a wide range of applications and domains integrating humans, computers, and knowledge.
Code (0)
등록된 구현이 없습니다.
Tasks
AutoMLExplainable artificial intelligenceExplainable Artificial Intelligence (XAI)Knowledge GraphsLanguage ModelingLanguage ModellingNatural Language UnderstandingSimilar Papers 제목 키워드 기반
PalimpChat: Declarative and Interactive AI analytics
Thanks to the advances in generative architectures and large language models, data scientists can now code pipelines of machine-learning operations to process large collections of unstructured data. Recent progress has s…
scientific discoverySemPiper: Interactive Code Synthesis for Semantic Operators in Machine Learning Pipelines
Machine learning (ML) pipelines require extensive data preparation, feature engineering, and integration across heterogeneous sources, making them tedious and error-prone to develop. While large language models (LLMs) ha…
Feature EngineeringAI for Low-Code for AI
Low-code programming allows citizen developers to create programs with minimal coding effort, typically via visual (e.g. drag-and-drop) interfaces. In parallel, recent AI-powered tools such as Copilot and ChatGPT generat…
BENTO: A Visual Platform for Building Clinical NLP Pipelines Based on CodaLab
CodaLab is an open-source web-based platform for collaborative computational research. Although CodaLab has gained popularity in the research community, its interface has limited support for creating reusable tools that …
Managementnamed-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)Correction and Corruption: A Two-Rate View of Error Flow in LLM Protocols
Large language models are increasingly deployed as protocols: structured multi-call procedures that spend additional computation to transform a baseline answer into a final one. These protocols are evaluated only by end-…