paper-with-me

홈 › Papers

A New Broad NLP Training from Speech to Knowledge

2021-06-01 · NAACL (TeachingNLP) 2021 6 · Maxime Amblard, Miguel Couceiro

In 2018, the Master Sc. in NLP opened at IDMC - Institut des Sciences du Digital, du Management et de la Cognition, Université de Lorraine - Nancy, France. Far from being a creation ex-nihilo, it is the product of a history and many reflections on the field and its teaching. This article proposes epistemological and critical elements on the opening and maintainance of this so far new master’s program in NLP.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Management

Similar Papers 제목 키워드 기반

Raon-Speech Technical Report

2026-04-08 · Beomsoo Kim, Changho Choi, Dohyun Kim, Dongki Lee 외 arxiv

We present Raon-Speech, a top-performing 9B-parameter speech language model (SpeechLM) for English and Korean speech understanding, answering, and generation, and Raon-SpeechChat, a high-performing full-duplex extension …

Knowledge DistillationQuestion Answering

DRAFT: A Novel Framework to Reduce Domain Shifting in Self-supervised Learning and Its Application to Children's ASR

2022-06-16 · Ruchao Fan, Abeer Alwan

Self-supervised learning (SSL) in the pretraining stage using un-annotated speech data has been successful in low-resource automatic speech recognition (ASR) tasks. However, models trained through SSL are biased to the p…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Self-Supervised Learningspeech-recognition+2

Dialog+ in Broadcasting: First Field Tests Using Deep-Learning-Based Dialogue Enhancement

2021-12-17 · Matteo Torcoli, Christian Simon, Jouni Paulus, Davide Straninger 외

Difficulties in following speech due to loud background sounds are common in broadcasting. Object-based audio, e.g., MPEG-H Audio solves this problem by providing a user-adjustable speech level. While object-based audio …

Object

Exploiting the large-scale German Broadcast Corpus to boost the Fraunhofer IAIS Speech Recognition System

2014-05-01 · LREC 2014 5 · Michael Stadtschnitzer, Jochen Schwenninger, Daniel Stein, Joachim Koehler

In this paper we describe the large-scale German broadcast corpus (GER-TV1000h) containing more than 1,000 hours of transcribed speech data. This corpus is unique in the German language corpora domain and enables signifi…

Acoustic ModellingAutomatic Speech RecognitionAutomatic Speech Recognition (ASR)Language Modelling+5

Two-Step Knowledge Distillation for Tiny Speech Enhancement

2023-09-15 · Rayan Daod Nathoo, Mikolaj Kegler, Marko Stamenovic

Tiny, causal models are crucial for embedded audio machine learning applications. Model compression can be achieved via distilling knowledge from a large teacher into a smaller student model. In this work, we propose a n…

Knowledge DistillationModel CompressionSpeech Enhancement