paper-with-me

Papers Decipherment

“Decipherment” 태그가 달린 논문 46편 · 필터 해제

OracleFusion: Assisting the Decipherment of Oracle Bone Script with Structurally Constrained Semantic Typography

2025-06-26 · Caoshuo Li, Zengmao Ding, Xiaobin Hu, Bang Li 외

As one of the earliest ancient languages, Oracle Bone Script (OBS) encapsulates the cultural records and intellectual expressions of ancient civilizations. Despite the discovery of approximately 4,500 OBS characters, onl…

DeciphermentLarge Language ModelMultimodal Large Language ModelVisual Localization

Ancient Script Image Recognition and Processing: A Review

2025-06-24 · Xiaolei Diao, Rite Bo, Yanling Xiao, Lida Shi 외

Ancient scripts, e.g., Egyptian hieroglyphs, Oracle Bone Inscriptions, and Ancient Greek inscriptions, serve as vital carriers of human civilization, embedding invaluable historical and cultural information. Automating a…

DeciphermentFew-Shot Learning

Reasoning Over the Glyphs: Evaluation of LLM's Decipherment of Rare Scripts

2025-01-29 · Yu-Fei Shih, Zheng-Lin Lin, Shu-Kai Hsieh

We explore the capabilities of LVLMs and LLMs in deciphering rare scripts not encoded in Unicode. We introduce a novel approach to construct a multimodal dataset of linguistic puzzles involving such scripts, utilizing a …

Decipherment

Determination of language families using deep learning

2024-09-04 · Peter B. Lerner

We use a c-GAN (convolutional generative adversarial) neural network to analyze transliterated text fragments of extant, dead comprehensible, and one dead non-deciphered (Cypro-Minoan) language to establish linguistic af…

DeciphermentDeep LearningTranslation

Oracle Bone Inscriptions Multi-modal Dataset

2024-07-04 · Bang Li, Donghao Luo, Yujie Liang, Jing Yang 외

Oracle bone inscriptions(OBI) is the earliest developed writing system in China, bearing invaluable written exemplifications of early Shang history and paleography. However, the task of deciphering OBI, in the current cl…

DeciphermentDenoising

Decipherment-Aware Multilingual Learning in Jointly Trained Language Models

2024-06-11 · Grandee Lee

The principle that governs unsupervised multilingual learning (UCL) in jointly trained language models (mBERT as a popular example) is still being debated. Many find it surprising that one can achieve UCL with multiple m…

Decipherment

Deciphering Oracle Bone Language with Diffusion Models

2024-06-02 · Haisu Guan, Huanxin Yang, Xinyu Wang, Shengwei Han 외

Originating from China's Shang Dynasty approximately 3,000 years ago, the Oracle Bone Script (OBS) is a cornerstone in the annals of linguistic history, predating many established writing systems. Despite the discovery o…

DeciphermentImage Generation

Segmentation of Maya hieroglyphs through fine-tuned foundation models

2024-05-26 · FNU Shivam, Megan Leight, Mary Kate Kelly, Claire Davis 외

The study of Maya hieroglyphic writing unlocks the rich history of cultural and societal knowledge embedded within this ancient civilization's visual narrative. Artificial Intelligence (AI) offers a novel lens through wh…

Decipherment

An open dataset for oracle bone script recognition and decipherment

2024-01-27 · Pengjie Wang, Kaile Zhang, Xinyu Wang, Shengwei Han 외

Oracle bone script, one of the earliest known forms of ancient Chinese writing, presents invaluable research materials for scholars studying the humanities and geography of the Shang Dynasty, dating back 3,000 years. The…

Decipherment

An open dataset for the evolution of oracle bone characters: EVOBC

2024-01-23 · Haisu Guan, Jinpeng Wan, Yuliang Liu, Pengjie Wang 외

The earliest extant Chinese characters originate from oracle bone inscriptions, which are closely related to other East Asian languages. These inscriptions hold immense value for anthropology and archaeology. However, de…

Decipherment

ChatABL: Abductive Learning via Natural Language Interaction with ChatGPT

2023-04-21 · Tianyang Zhong, Yaonai Wei, Li Yang, Zihao Wu 외

Large language models (LLMs) such as ChatGPT have recently demonstrated significant potential in mathematical abilities, providing valuable reasoning paradigm consistent with human natural language. However, LLMs current…

DeciphermentLogical Reasoning

From Inscription to Semi-automatic Annotation of Maya Hieroglyphic Texts

2022-06-01 · LT4HALA (LREC) 2022 6 · Cristina Vertan, Christian Prager

The Maya script is the only readable autochthonous writing system of the Americas and consists of more than 1000 word signs and syllables. It is only partially deciphered and is the subject of the project “Text Database …

DeciphermentTransliteration

Dorabella Cipher as Musical Inspiration

2021-12-01 · SMP (ICON) 2021 12 · Bradley Hauer, Colin Choi, Abram Hindle, Scott Smallwood 외

The Dorabella cipher is an encrypted note of English composer Edward Elgar, which has defied decipherment attempts for more than a century. While most proposed solutions are English texts, we investigate the hypothe- sis…

DeciphermentPosition

Deciphering Speech: a Zero-Resource Approach to Cross-Lingual Transfer in ASR

2021-11-12 · Ondrej Klejch, Electra Wallington, Peter Bell

We present a method for cross-lingual training an ASR system using absolutely no transcribed training data from the target language, and with no phonetic knowledge of the language in question. Our approach uses a novel a…

Cross-Lingual ASRCross-Lingual TransferDecipherment

A Text GAN for Language Generation with Non-Autoregressive Generator

2021-01-01 · Fei Huang, Jian Guan, Pei Ke, Qihan Guo 외

Despite the great success of Generative Adversarial Networks (GANs) in generating high-quality images, GANs for text generation still face two major challenges: first, most text GANs are unstable in training mainly due t…

DeciphermentRepresentation LearningSentenceText Generation

Can Sequence-to-Sequence Models Crack Substitution Ciphers?

2020-12-30 · ACL 2021 5 · Nada Aldarrab, Jonathan May

Decipherment of historical ciphers is a challenging problem. The language of the target plaintext might be unknown, and ciphertext can have a lot of noise. State-of-the-art decipherment methods use beam search and a neur…

DeciphermentLanguage IdentificationLanguage ModelingLanguage Modelling

Deciphering Undersegmented Ancient Scripts Using Phonetic Prior

2020-10-21 · Jiaming Luo, Frederik Hartmann, Enrico Santus, Yuan Cao 외

Most undeciphered lost languages exhibit two characteristics that pose significant decipherment challenges: (1) the scripts are not fully segmented into words; (2) the closest known language is not determined. We propose…

Decipherment

Phonetic and Visual Priors for Decipherment of Informal Romanization

2020-05-05 · ACL 2020 6 · Maria Ryskina, Matthew R. Gormley, Taylor Berg-Kirkpatrick

Informal romanization is an idiosyncratic process used by humans in informal digital communication to encode non-Latin script languages into Latin character sets found on common keyboards. Character substitution choices …

DeciphermentInductive Bias

A Probabilistic Formulation of Unsupervised Text Style Transfer

2020-02-10 · ICLR 2020 1 · Junxian He, Xinyi Wang, Graham Neubig, Taylor Berg-Kirkpatrick

We present a deep generative model for unsupervised text style transfer that unifies previously proposed non-generative techniques. Our probabilistic approach models non-parallel data from two domains as a partially obse…

DeciphermentLanguage ModellingMachine TranslationStyle Transfer+5

Neural Decipherment via Minimum-Cost Flow: from Ugaritic to Linear B

2019-06-16 · ACL 2019 7 · Jiaming Luo, Yuan Cao, Regina Barzilay

In this paper we propose a novel neural approach for automatic decipherment of lost languages. To compensate for the lack of strong supervision signal, our model design is informed by patterns in language change document…

Decipherment
1–20 / 46 다음 →