paper-with-me

Papers

A frame semantics based approach to comparative study of digitized corpus

2020-05-29 · Abdelaziz Lakhfif, Mohamed Tayeb Laskri

in this paper, we present a corpus linguistics based approach applied to analyzing digitized classical multilingual novels and narrative texts, from a semantic point of view. Digitized novels such as "the hobbit (Tolkien J. R. R., 1937)" and "the hound of the Baskervilles (Doyle A. C. 1901-1902)", which were widely translated to dozens of languages, provide rich materials for analyzing languages differences from several perspectives and within a number of disciplines like linguistics, philosophy and cognitive science. Taking motion events conceptualization as a case study, this paper, focus on the morphologic, syntactic, and semantic annotation process of English-Arabic aligned corpus created from a digitized novels, in order to re-examine the linguistic encodings of motion events in English and Arabic in terms of Frame Semantics. The present study argues that differences in motion events conceptualization across languages can be described with frame structure and frame-to-frame relations.

📄 PDF Abstract BibTeX arXiv:2006.00113

Code (0)

등록된 구현이 없습니다.

Tasks

Philosophy

Similar Papers 제목 키워드 기반

The development of a web corpus of Hindi language and corpus-based comparative studies to Japanese

2016-12-01 · WS 2016 12 · Miki Nishioka, Shiro Akasegawa

In this paper, we discuss our creation of a web corpus of spoken Hindi (COSH), one of the Indo-Aryan languages spoken mainly in the Indian subcontinent. We also point out notable problems we{'}ve encountered in the web c…

The DReaM Corpus: A Multilingual Annotated Corpus of Grammars for the World's Languages

2020-05-01 · LREC 2020 5 · Shafqat Mumtaz Virk, Harald Hammarstr{\"o}m, Markus Forsberg, S{\o}ren Wichmann

There exist as many as 7000 natural languages in the world, and a huge number of documents describing those languages have been produced over the years. Most of those documents are in paper format. Any attempts to use mo…

Correcting Whitespace Errors in Digitized Historical Texts

2019-06-01 · WS 2019 6 · S Soni, eep, Lauren Klein, Jacob Eisenstein

Whitespace errors are common to digitized archives. This paper describes a lightweight unsupervised technique for recovering the original whitespace. Our approach is based on count statistics from Google n-grams, which a…

A CCG-based Compositional Semantics and Inference System for Comparatives

2019-10-02 · Izumi Haruta, Koji Mineshima, Daisuke Bekki

Comparative constructions play an important role in natural language inference. However, attempts to study semantic representations and logical inferences for comparatives from the computational perspective are not well …

Natural Language Inference

Can Large Language Models (LLMs) Describe Pictures Like Children? A Comparative Corpus Study

2025-08-19 · Hanna Woloszyn, Benjamin Gagl arxiv

The role of large language models (LLMs) in education is increasing, yet little attention has been paid to whether LLM-generated text resembles child language. This study evaluates how LLMs replicate child-like language …

Semantic Similarity