paper-with-me

홈 › Papers

DomainCQA: Crafting Expert-Level QA from Domain-Specific Charts

2025-03-25 · Ling Zhong, Yujing Lu, Jing Yang, Weiming Li, Peng Wei, Yongheng Wang, Manni Duan, Qing Zhang

Chart Question Answering (CQA) benchmarks are essential for evaluating the capability of Multimodal Large Language Models (MLLMs) to interpret visual data. However, current benchmarks focus primarily on the evaluation of general-purpose CQA but fail to adequately capture domain-specific challenges. We introduce DomainCQA, a systematic methodology for constructing domain-specific CQA benchmarks, and demonstrate its effectiveness by developing AstroChart, a CQA benchmark in the field of astronomy. Our evaluation shows that chart reasoning and combining chart information with domain knowledge for deeper analysis and summarization, rather than domain-specific knowledge, pose the primary challenge for existing MLLMs, highlighting a critical gap in current benchmarks. By providing a scalable and rigorous framework, DomainCQA enables more precise assessment and improvement of MLLMs for domain-specific applications.

📄 PDF Abstract BibTeX arXiv:2503.19498

Code (0)

등록된 구현이 없습니다.

Tasks

AstronomyChart Question AnsweringQuestion Answering

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

I Know Therefore I Score: Label-Free Crafting of Scoring Functions using Constraints Based on Domain Expertise

2022-03-18 · Ragja Palakkadavath, Sarath Sivaprasad, Shirish Karande, Niranjan Pedanekar

Several real-life applications require crafting concise, quantitative scoring functions (also called rating systems) from measured observations. For example, an effectiveness score needs to be created for advertising cam…

Specific Domain Ontology Construction Using Large Language Models

2026-06-14 · Vivian Magri Alcaldi Soares, Renata Wassermann arxiv

Ontologies are useful structures to organize and maintain information that can be understood both by humans and systems. However, since their manual crafting is a laborious task, many specific domains lack reference onto…

CHiLL: Zero-shot Custom Interpretable Feature Extraction from Clinical Notes with Large Language Models

2023-02-23 · Denis Jered McInerney, Geoffrey Young, Jan-Willem van de Meent, Byron C. Wallace

We propose CHiLL (Crafting High-Level Latents), an approach for natural-language specification of features for linear models. CHiLL prompts LLMs with expert-crafted queries to generate interpretable features from health …

Dynamic Context-Aware Prompt Recommendation for Domain-Specific AI Applications

2025-06-25 · Xinye Tang, Haijun Zhai, Chaitanya Belwal, Vineeth Thayanithi 외

LLM-powered applications are highly susceptible to the quality of user prompts, and crafting high-quality prompts can often be challenging especially for domain-specific applications. This paper presents a novel dynamic …

Few-Shot Learning

ELF-Gym: Evaluating Large Language Models Generated Features for Tabular Prediction

2024-10-13 · Yanlin Zhang, Ning li, Quan Gan, Weinan Zhang 외

Crafting effective features is a crucial yet labor-intensive and domain-specific task within machine learning pipelines. Fortunately, recent advancements in Large Language Models (LLMs) have shown promise in automating v…

Feature Engineering