paper-with-me

홈 › Papers

ATHAR: A High-Quality and Diverse Dataset for Classical Arabic to English Translation

2024-07-29 · Mohammed Khalil, Mohammed Sabry

Classical Arabic represents a significant era, encompassing the golden age of Arab culture, philosophy, and scientific literature. With a broad consensus on the importance of translating these literatures to enrich knowledge dissemination across communities, the advent of large language models (LLMs) and translation systems offers promising tools to facilitate this goal. However, we have identified a scarcity of translation datasets in Classical Arabic, which are often limited in scope and topics, hindering the development of high-quality translation systems. In response, we present the ATHAR dataset, comprising 66,000 high-quality Classical Arabic to English translation samples that cover a wide array of subjects including science, culture, and philosophy. Furthermore, we assess the performance of current state-of-the-art LLMs under various settings, concluding that there is a need for such datasets in current systems. Our findings highlight how models can benefit from fine-tuning or incorporating this dataset into their pretraining pipelines. The dataset is publicly available on the HuggingFace Data Hub at \url{https://huggingface.co/datasets/mohamed-khalil/ATHAR}.

📄 PDF Abstract BibTeX arXiv:2407.19835

Code (0)

등록된 구현이 없습니다.

Tasks

PhilosophyTranslation

Similar Papers 제목 키워드 기반

MathArena: Evaluating LLMs on Uncontaminated Math Competitions

2025-05-29 · Mislav Balunović, Jasper Dekoninck, Ivo Petrov, Nikola Jovanović 외

The rapid advancement of reasoning capabilities in large language models (LLMs) has led to notable improvements on mathematical benchmarks. However, many of the most commonly used evaluation datasets (e.g., AIME 2024) ar…

MathMathematical ReasoningMemorization

Beyond Benchmarks: MathArena as an Evaluation Platform for Mathematics with LLMs

2026-05-01 · Jasper Dekoninck, Nikola Jovanović, Tim Gehrunger, Kári Rögnvaldsson 외 arxiv

Large language models (LLMs) are becoming increasingly capable mathematical collaborators, but static benchmarks are no longer sufficient for evaluating progress: they are often narrow in scope, quickly saturated, and ra…

Mathematical Reasoning

PathAR: Structure-First Autoregressive Synthesis of Multimodal Pathology Images

2026-06-01 · Yuan Zhang, Jiahao Xia, Junzhang Huang, Meng Wang 외 arxiv

Data scarcity in multimodal pathology motivates unified generative models that synthesize modality-specific appearance while preserving anatomically coherent structure. Although modalities differ in appearance statistics…

ChatHaruhi: Reviving Anime Character in Reality via Large Language Model

2023-08-18 · Cheng Li, Ziang Leng, Chenxi Yan, Junyi Shen 외

Role-playing chatbots built on large language models have drawn interest, but better techniques are needed to enable mimicking specific fictional characters. We propose an algorithm that controls language models via an i…

Language ModelingLanguage ModellingLarge Language ModelText2text Generation+1

Multitask Learning Can Improve Worst-Group Outcomes

2023-12-05 · Atharva Kulkarni, Lucio Dery, Amrith Setlur, aditi raghunathan 외

In order to create machine learning systems that serve a variety of users well, it is vital to not only achieve high average performance but also ensure equitable outcomes across diverse groups. However, most machine lea…

Fairness