paper-with-me

Papers

Exploring Unexplored Generalization Challenges for Cross-Database Semantic Parsing

2020-07-01 · ACL 2020 6 · Alane Suhr, Ming-Wei Chang, Peter Shaw, Kenton Lee

We study the task of cross-database semantic parsing (XSP), where a system that maps natural language utterances to executable SQL queries is evaluated on databases unseen during training. Recently, several datasets, including Spider, were proposed to support development of XSP systems. We propose a challenging evaluation setup for cross-database semantic parsing, focusing on variation across database schemas and in-domain language use. We re-purpose eight semantic parsing datasets that have been well-studied in the setting where in-domain training data is available, and instead use them as additional evaluation data for XSP systems instead. We build a system that performs well on Spider, and find that it struggles to generalize to our re-purposed set. Our setup uncovers several generalization challenges for cross-database semantic parsing, demonstrating the need to use and develop diverse training and evaluation datasets.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Semantic Parsing

Similar Papers 제목 키워드 기반

Defining and Counting Phonological Classes in Cross-linguistic Segment Databases

2016-05-01 · LREC 2016 5 · Dan Dediu, Scott Moisik

Recently, there has been an explosion in the availability of large, good-quality cross-linguistic databases such as WALS (Dryer {\&} Haspelmath, 2013), Glottolog (Hammarstrom et al., 2015) and Phoible (Moran {\&} McCloy,…

Diversity

On the Structural Generalization in Text-to-SQL

2023-01-12 · Jieyu Li, Lu Chen, Ruisheng Cao, Su Zhu 외

Exploring the generalization of a text-to-SQL parser is essential for a system to automatically adapt the real-world databases. Previous works provided investigations focusing on lexical diversity, including the influenc…

DiversityText to SQLText-To-SQL

Tackling prediction tasks in relational databases with LLMs

2024-11-18 · Marek Wydmuch, Łukasz Borchmann, Filip Graliński

Though large language models (LLMs) have demonstrated exceptional performance across numerous problems, their application to predictive tasks in relational databases remains largely unexplored. In this work, we address t…

Prediction

Exploring Scaling Laws for EHR Foundation Models

2025-05-29 · Sheng Zhang, Qin Liu, Naoto Usuyama, Cliff Wong 외

The emergence of scaling laws has profoundly shaped the development of large language models (LLMs), enabling predictable performance gains through systematic increases in model size, dataset volume, and compute. Yet, th…

Impact of ECG Dataset Diversity on Generalization of CNN Model for Detecting QRS Complex

2019-07-10 · IEEE Access 2019 7 · Ahsan Habib, Chandan Karmakar, John Yearwood

Detection of QRS complexes in electrocardiogram (ECG) signal is crucial for automated cardiac diagnosis. Automated QRS detection has been a research topic for over three decades and several of the traditional QRS detecti…

DiversityElectrocardiography (ECG)QRS Complex Detection