paper-with-me

홈 › Papers

Spanish Built Factual Freectianary (Spanish-BFF): the first AI-generated free dictionary

2023-02-24 · Miguel Ortega-Martín, Óscar García-Sierra, Alfonso Ardoiz, Juan Carlos Armenteros, Jorge Álvarez, Adrián Alonso

Dictionaries are one of the oldest and most used linguistic resources. Building them is a complex task that, to the best of our knowledge, has yet to be explored with generative Large Language Models (LLMs). We introduce the "Spanish Built Factual Freectianary" (Spanish-BFF) as the first Spanish AI-generated dictionary. This first-of-its-kind free dictionary uses GPT-3. We also define future steps we aim to follow to improve this initial commitment to the field, such as more additional languages.

📄 PDF Abstract BibTeX arXiv:2302.12746

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

{Dispute@FaQ-s}How to file a dispute with Expedia? How to file a dispute with Expedia? To file a complaint against Expedia, first try contacting their customer service directly. You can reach them by phone at…
Multi-Head Attention 설명 없음
Attention 설명 없음
Cosine Annealing Cosine Annealing is a type of learning rate schedule that has the effect of starting with a large learning rate that is relatively rapidly decreased to a minimum value before…
Linear Layer A Linear Layer is a projection $\mathbf{XW + b}$.
Dropout Dropout is a regularization technique for neural networks that drops a unit (along with connections) at training time with a specified probability $p$ (a common value is…
BPE Byte Pair Encoding, or BPE, is a subword segmentation algorithm that encodes rare and unknown words as sequences of subword units. The intuition is that various word…
Weight Decay 설명 없음

Similar Papers 제목 키워드 기반

Building another Spanish dictionary, this time with GPT-4

2024-06-17 · Miguel Ortega-Martín, Óscar García-Sierra, Alfonso Ardoiz, Juan Carlos Armenteros 외

We present the "Spanish Built Factual Freectianary 2.0" (Spanish-BFF-2) as the second iteration of an AI-generated Spanish dictionary. Previously, we developed the inaugural version of this unique free dictionary employi…

Free/Open Source Shallow-Transfer Based Machine Translation for Spanish and Aragonese

2012-05-01 · LREC 2012 5 · Juan Pablo Mart{\'\i}nez Cort{\'e}s, Jim O{'}Regan, Francis Tyers

This article describes the development of a bidirectional shallow-transfer based machine translation system for Spanish and Aragonese, based on the Apertium platform, reusing the resources provided by other translators b…

Machine TranslationMorphological AnalysisTranslation

Factuality Annotation and Learning in Spanish Texts

2016-05-01 · LREC 2016 5 · Dina Wonsever, Aiala Ros{\'a}, Marisa Malcuori

We present a proposal for the annotation of factuality of event mentions in Spanish texts and a free available annotated corpus. Our factuality model aims to capture a pragmatic notion of factuality, trying to reflect a …

ACTIV-ES: a comparable, cross-dialect corpus of `everyday' Spanish from Argentina, Mexico, and Spain

2014-05-01 · LREC 2014 5 · Jerid Francom, Mans Hulden, Adam Ussishkin

Corpus resources for Spanish have proved invaluable for a number of applications in a wide variety of fields. However, a majority of resources are based on formal, written language and/or are not built to model language …

Part-Of-Speech Tagging

Crowdsourcing Latin American Spanish for Low-Resource Text-to-Speech

2020-05-01 · LREC 2020 5 · Adriana Guevara-Rukoz, Isin Demirsahin, Fei He, Shan-Hui Cathy Chu 외

In this paper we present a multidialectal corpus approach for building a text-to-speech voice for a new dialect in a language with existing resources, focusing on various South American dialects of Spanish. We first pres…

text-to-speechText to Speech