paper-with-me

Papers

Evaluation Ethics of LLMs in Legal Domain

2024-03-17 · Ruizhe Zhang, Haitao Li, Yueyue Wu, Qingyao Ai, Yiqun Liu, Min Zhang, Shaoping Ma

In recent years, the utilization of large language models for natural language dialogue has gained momentum, leading to their widespread adoption across various domains. However, their universal competence in addressing challenges specific to specialized fields such as law remains a subject of scrutiny. The incorporation of legal ethics into the model has been overlooked by researchers. We asserts that rigorous ethic evaluation is essential to ensure the effective integration of large language models in legal domains, emphasizing the need to assess domain-specific proficiency and domain-specific ethic. To address this, we propose a novelty evaluation methodology, utilizing authentic legal cases to evaluate the fundamental language abilities, specialized legal knowledge and legal robustness of large language models (LLMs). The findings from our comprehensive evaluation contribute significantly to the academic discourse surrounding the suitability and performance of large language models in legal domains.

📄 PDF Abstract BibTeX arXiv:2403.11152

Code (0)

등록된 구현이 없습니다.

Tasks

Ethics

Similar Papers 제목 키워드 기반

A Human-Centric Pipeline for Aligning Large Language Models with Chinese Medical Ethics

2026-01-12 · Haoan Jin, Han Ying, Jiacheng Ji, Hanhui Xu 외 arxiv

Recent advances in large language models have enabled their application to a range of healthcare tasks. However, aligning LLMs with the nuanced demands of medical ethics, especially under complex real world scenarios, re…

TRIDENT: Benchmarking LLM Safety in Finance, Medicine, and Law

2025-07-22 · Zheng Hui, Yijiang River Dong, Ehsan Shareghi, Nigel Collier arxiv

As large language models (LLMs) are increasingly deployed in high-risk domains such as law, finance, and medicine, systematically evaluating their domain-specific safety and compliance becomes critical. While prior work …

LawBench: Benchmarking Legal Knowledge of Large Language Models

2023-09-28 · Zhiwei Fei, Xiaoyu Shen, Dawei Zhu, Fengzhe Zhou 외

Large language models (LLMs) have demonstrated strong capabilities in various aspects. However, when applying them to the highly specialized, safe-critical legal domain, it is unclear how much legal knowledge they posses…

ArticlesBenchmarkingMemorizationMulti-Label Classification+1

LAiW: A Chinese Legal Large Language Models Benchmark

2023-10-09 · Yongfu Dai, Duanyu Feng, Jimin Huang, Haochen Jia 외

General and legal domain LLMs have demonstrated strong performance in various tasks of LegalAI. However, the current evaluations of these LLMs in LegalAI are defined by the experts of computer science, lacking consistenc…

Information Retrieval

Evaluation of Large Language Models in Legal Applications: Challenges, Methods, and Future Directions

2026-01-21 · Yiran Hu, Huanghai Liu, Chong Wang, Kunran Li 외 arxiv

Large language models (LLMs) are being increasingly integrated into legal applications, including judicial decision support, legal practice assistance, and public-facing legal services. While LLMs show strong potential i…

Legal Reasoning