paper-with-me

홈 › Papers

Design principles of an open-source language modeling microservice package for AAC text-entry applications

2022-05-01 · SLPAT (ACL) 2022 5 · Brian Roark, Alexander Gutkin

We present MozoLM, an open-source language model microservice package intended for use in AAC text-entry applications, with a particular focus on the design principles of the library. The intent of the library is to allow the ensembling of multiple diverse language models without requiring the clients (user interface designers, system users or speech-language pathologists) to attend to the formats of the models. Issues around privacy, security, dynamic versus static models, and methods of model combination are explored and specific design choices motivated. Some simulation experiments demonstrating the benefits of personalized language model ensembling via the library are presented.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Language ModelingLanguage Modelling

Similar Papers 제목 키워드 기반

How to be FAIR when you CARE: The DGS Corpus as a Case Study of Open Science Resources for Minority Languages

2022-06-01 · LREC 2022 6 · Marc Schulder, Thomas Hanke

The publication of resources for minority languages requires a balance between making data open and accessible and respecting the rights and needs of its language community. The FAIR principles were introduced as a guide…

Cross-Layer Misalignment Detection in Agent Skills: A Progressive Loading-Aware Contrastive Learning Approach

2026-07-12 · Chengjun Zhang, Yang Gao, Jianna Hur, Jingjing Zhang 외 arxiv

Large language model (LLM) agents are increasingly extended through Agent Skills, reusable artifacts that package natural-language metadata, procedural instructions, and execution-time resources for runtime use. As open-…

Contrastive Learning

Dolma: an Open Corpus of Three Trillion Tokens for Language Model Pretraining Research

2024-01-31 · Luca Soldaini, Rodney Kinney, Akshita Bhagia, Dustin Schwenk 외

Information about pretraining corpora used to train the current best-performing language models is seldom discussed: commercial models rarely detail their data, and even open models are often released without accompanyin…

Language ModelingLanguage Modelling

CodeTF: One-stop Transformer Library for State-of-the-art Code LLM

2023-05-31 · Nghi D. Q. Bui, Hung Le, Yue Wang, Junnan Li 외

Code intelligence plays a key role in transforming modern software engineering. Recently, deep learning-based models, especially Transformer-based large language models (LLMs), have demonstrated remarkable potential in t…

Adapting Language Specific Components of Cross-Media Analysis Frameworks to Less-Resourced Languages: the Case of Amharic

2020-05-01 · LREC 2020 5 · Yonas Woldemariam, Adam Dahlgren

We present an ASR based pipeline for Amharic that orchestrates NLP components within a cross media analysis framework (CMAF). One of the major challenges that are inherently associated with CMAFs is effectively addressin…

named-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)NER