paper-with-me

Papers

PhiloBERTA: A Transformer-Based Cross-Lingual Analysis of Greek and Latin Lexicons

2025-03-07 · Rumi A. Allbert, Makai L. Allbert

We present PhiloBERTA, a cross-lingual transformer model that measures semantic relationships between ancient Greek and Latin lexicons. Through analysis of selected term pairs from classical texts, we use contextual embeddings and angular similarity metrics to identify precise semantic alignments. Our results show that etymologically related pairs demonstrate significantly higher similarity scores, particularly for abstract philosophical concepts such as epist\=em\=e (scientia) and dikaiosyn\=e (iustitia). Statistical analysis reveals consistent patterns in these relationships (p = 0.012), with etymologically related pairs showing remarkably stable semantic preservation compared to control pairs. These findings establish a quantitative framework for examining how philosophical concepts moved between Greek and Latin traditions, offering new methods for classical philological research.

📄 PDF Abstract BibTeX arXiv:2503.05265

Code (1)

RumiAllbert/PhiloBERTA 공식 구현 pytorch

Similar Papers 제목 키워드 기반

A Comparative Evaluation of Embeddings and LLMs in a Greek Book Publisher Setting - The CUP Dataset

2026-07-23 · Katerina Papantoniou, Panagiotis Papadakos, Theodore Patkos, Dimitris Garefalakis 외 arxiv

We present CUP, a Greek book retrieval benchmark consisting of 868 catalog records and 104 expert-annotated queries with graded relevance judgments. We evaluate sparse (BM25), dense (sentence-transformers), hybrid, and L…

Multi-granular Legal Topic Classification on Greek Legislation

2021-09-30 · EMNLP (NLLP) 2021 11 · Christos Papaloukas, Ilias Chalkidis, Konstantinos Athinaios, Despina-Athanasia Pantazi 외

In this work, we study the task of classifying legal texts written in the Greek language. We introduce and make publicly available a novel dataset based on Greek legislation, consisting of more than 47 thousand official,…

Classificationtext-classificationText ClassificationTopic Classification+2

Forging GEMs: Advancing Greek NLP through Quality-Based Corpus Curation

2025-10-22 · Alexandra Apostolopoulou, Konstantinos Kanaris, Athanasios Koursaris, Dimitris Tsakalidis 외 arxiv

The advancement of natural language processing for morphologically rich and moderately-resourced languages like Modern Greek has been hindered by architectural stagnation, data scarcity, and limited context processing ca…

Natural Language UnderstandingDomain Adaptation

GREEK-BERT: The Greeks visiting Sesame Street

2020-08-27 · John Koutsikakis, Ilias Chalkidis, Prodromos Malakasiotis, Ion Androutsopoulos

Transformer-based language models, such as BERT and its variants, have achieved state-of-the-art performance in several downstream natural language processing (NLP) tasks on generic benchmark datasets (e.g., GLUE, SQUAD,…

Language ModelingLanguage Modellingnamed-entity-recognitionNamed Entity Recognition+5

Analysing the Greek Parliament Records with Emotion Classification

2022-05-24 · John Pavlopoulos, Vanessa Lislevand

In this project, we tackle emotion classification for the Greek language, presenting and releasing a new dataset in Greek. We fine-tune and assess Transformer-based masked language models that were pre-trained on monolin…

ClassificationEmotion Classification