paper-with-me

홈 › Papers

VLQA: The First Comprehensive, Large, and High-Quality Vietnamese Dataset for Legal Question Answering

2025-07-26 · Tan-Minh Nguyen, Hoang-Trung Nguyen, Trong-Khoi Dao, Xuan-Hieu Phan, Ha-Thanh Nguyen, Thi-Hai-Yen Vuong arxiv

The advent of large language models (LLMs) has led to significant achievements in various domains, including legal text processing. Leveraging LLMs for legal tasks is a natural evolution and an increasingly compelling choice. However, their capabilities are often portrayed as greater than they truly are. Despite the progress, we are still far from the ultimate goal of fully automating legal tasks using artificial intelligence (AI) and natural language processing (NLP). Moreover, legal systems are deeply domain-specific and exhibit substantial variation across different countries and languages. The need for building legal text processing applications for different natural languages is, therefore, large and urgent. However, there is a big challenge for legal NLP in low-resource languages such as Vietnamese due to the scarcity of resources and annotated data. The need for labeled legal corpora for supervised training, validation, and supervised fine-tuning is critical. In this paper, we introduce the VLQA dataset, a comprehensive and high-quality resource tailored for the Vietnamese legal domain. We also conduct a comprehensive statistical analysis of the dataset and evaluate its effectiveness through experiments with state-of-the-art models on legal information retrieval and question-answering tasks.

📄 PDF Abstract BibTeX arXiv:2507.19995

Code (0)

등록된 구현이 없습니다.

Tasks

Information RetrievalQuestion Answering

Similar Papers 제목 키워드 기반

Visuo-Linguistic Question Answering (VLQA) Challenge

2020-05-01 · Findings of the Association for Computational Linguistics 2020 · Shailaja Keyur Sampat, Yezhou Yang, Chitta Baral

Understanding images and text together is an important aspect of cognition and building advanced Artificial Intelligence (AI) systems. As a community, we have achieved good benchmarks over language and vision domains sep…

Question AnsweringReading ComprehensionVisual Question Answering (VQA)

GITA: Graph to Visual and Textual Integration for Vision-Language Graph Reasoning

2024-02-03 · Yanbin Wei, Shuai Fu, Weisen Jiang, Zejian Zhang 외

Large Language Models (LLMs) are increasingly used for various tasks with graph structures. Though LLMs can process graph information in a textual format, they overlook the rich vision modality, which is an intuitive way…

Link PredictionNode Classification

Research on Vision-Language Question Answering Models for Industrial Robots

2026-05-02 · Ping Li, Bartlomiej Brzozka arxiv

A hierarchical cross-modal fusion model is proposed for vision-language question answering (VLQA) in industrial robotics, targeting the challenges of semantic ambiguity, complex environmental layouts, and domain-specific…

Question AnsweringAnomaly DetectionObject Detection

Towards Retrieval Augmented Generation over Large Video Libraries

2024-06-21 · Yannis Tevissen, Khalil Guetari, Frédéric Petitpont

Video content creators need efficient tools to repurpose content, a task that often requires complex manual or automated searches. Crafting a new video from large video libraries remains a challenge. In this paper we int…

Answer GenerationQuestion AnsweringRAGRetrieval+1

Chinese SimpleQA: A Chinese Factuality Evaluation for Large Language Models

2024-11-11 · Yancheng He, Shilong Li, Jiaheng Liu, Yingshui Tan 외

New LLM evaluation benchmarks are important to align with the rapid development of Large Language Models (LLMs). In this work, we present Chinese SimpleQA, the first comprehensive Chinese benchmark to evaluate the factua…