paper-with-me

홈 › Papers

Comprehensive Evaluation for a Large Scale Knowledge Graph Question Answering Service

2025-01-28 · Saloni Potdar, Daniel Lee, Omar Attia, Varun Embar, De Meng, Ramesh Balaji, Chloe Seivwright, Eric Choi, Mina H. Farid, Yiwen Sun, Yunyao Li

Question answering systems for knowledge graph (KGQA), answer factoid questions based on the data in the knowledge graph. KGQA systems are complex because the system has to understand the relations and entities in the knowledge-seeking natural language queries and map them to structured queries against the KG to answer them. In this paper, we introduce Chronos, a comprehensive evaluation framework for KGQA at industry scale. It is designed to evaluate such a multi-component system comprehensively, focusing on (1) end-to-end and component-level metrics, (2) scalable to diverse datasets and (3) a scalable approach to measure the performance of the system prior to release. In this paper, we discuss the unique challenges associated with evaluating KGQA systems at industry scale, review the design of Chronos, and how it addresses these challenges. We will demonstrate how it provides a base for data-driven decisions and discuss the challenges of using it to measure and improve a real-world KGQA system.

📄 PDF Abstract BibTeX arXiv:2501.17270

Code (0)

등록된 구현이 없습니다.

Tasks

Graph Question AnsweringNatural Language QueriesQuestion Answering

Methods 이 논문이 사용한 방법론

BASE 설명 없음

Similar Papers 제목 키워드 기반

PubGraph: A Large-Scale Scientific Knowledge Graph

2023-02-04 · Kian Ahrabian, Xinwei Du, Richard Delwin Myloth, Arun Baalaaji Sankar Ananthan 외

Research publications are the primary vehicle for sharing scientific progress in the form of new discoveries, methods, techniques, and insights. Unfortunately, the lack of a large-scale, comprehensive, and easy-to-use re…

Community DetectionGraph EmbeddingInductive LearningKnowledge Graph Completion+2

UniEdit: A Unified Knowledge Editing Benchmark for Large Language Models

2025-05-18 · Qizhou Chen, Dakan Wang, Taolin Zhang, Zaoming Yan 외

Model editing aims to enhance the accuracy and reliability of large language models (LLMs) by efficiently adjusting their internal parameters. Currently, most LLM editing datasets are confined to narrow knowledge domains…

Diversityknowledge editingKnowledge GraphsModel Editing

On Large-scale Evaluation of Embedding Models for Knowledge Graph Completion

2025-04-11 · Nasim Shirvani-Mahdavi, Farahnaz Akrami, Chengkai Li

Knowledge graph embedding (KGE) models are extensively studied for knowledge graph completion, yet their evaluation remains constrained by unrealistic benchmarks. Standard evaluation metrics rely on the closed-world assu…

Graph EmbeddingKnowledge Graph CompletionKnowledge Graph EmbeddingLink Prediction+2

TGB 2.0: A Benchmark for Learning on Temporal Knowledge Graphs and Heterogeneous Graphs

2024-06-14 · Julia Gastinger, Shenyang Huang, Mikhail Galkin, Erfan Loghmani 외

Multi-relational temporal graphs are powerful tools for modeling real-world data, capturing the evolving and interconnected nature of entities over time. Recently, many novel models are proposed for ML on such graphs int…

BenchmarkingKnowledge Graphs

AFEC: A Knowledge Graph Capturing Social Intelligence in Casual Conversations

2022-05-22 · Yubo Xie, Junze Li, Pearl Pu

This paper introduces AFEC, an automatically curated knowledge graph based on people's day-to-day casual conversations. The knowledge captured in this graph bears potential for conversational systems to understand how pe…

ChatbotDiversityRetrieval