paper-with-me

Papers

SBAAM! Eliminating Transcript Dependency in Automatic Subtitling

2024-05-17 · Marco Gaido, Sara Papi, Matteo Negri, Mauro Cettolo, Luisa Bentivogli

Subtitling plays a crucial role in enhancing the accessibility of audiovisual content and encompasses three primary subtasks: translating spoken dialogue, segmenting translations into concise textual units, and estimating timestamps that govern their on-screen duration. Past attempts to automate this process rely, to varying degrees, on automatic transcripts, employed diversely for the three subtasks. In response to the acknowledged limitations associated with this reliance on transcripts, recent research has shifted towards transcription-free solutions for translation and segmentation, leaving the direct generation of timestamps as uncharted territory. To fill this gap, we introduce the first direct model capable of producing automatic subtitles, entirely eliminating any dependence on intermediate transcripts also for timestamp prediction. Experimental results, backed by manual evaluation, showcase our solution's new state-of-the-art performance across multiple language pairs and diverse conditions.

📄 PDF Abstract BibTeX arXiv:2405.10741

Code (2)

hlt-mt/fbk-fairseq 공식 구현 pytorch
hlt-mt/subsonar 공식 구현 pytorch

Similar Papers 제목 키워드 기반

The PASSAGE project : Standard German Subtitling of Swiss German TV content

2022-06-01 · EAMT 2022 6 · Pierrette Bouillon, Johanna Gerlach, Jonathan Mutal, Marianne Starlander

We present the PASSAGE project, which aims at automatic Standard German subtitling of Swiss German TV content. This is achieved in a two step process, beginning with ASR to produce a normalised transcription, followed by…

Translation

MuST-Cinema: a Speech-to-Subtitles corpus

2020-02-25 · LREC 2020 5 · Alina Karakanta, Matteo Negri, Marco Turchi

Growing needs in localising audiovisual content in multiple languages through subtitles call for the development of automatic solutions for human subtitling. Neural Machine Translation (NMT) can contribute to the automat…

Machine TranslationNMTTranslation

Phoneme Similarity Matrices to Improve Long Audio Alignment for Automatic Subtitling

2014-05-01 · LREC 2014 5 · Pablo Ruiz, Aitor {\'A}lvarez, Haritz Arzelus

Long audio alignment systems for Spanish and English are presented, within an automatic subtitling application. Language-specific phone decoders automatically recognize audio contents at phoneme level. At the same time, …

DecoderLanguage Modelling

Advancing Arabic Speech Recognition Through Large-Scale Weakly Supervised Learning

2025-04-16 · Mahmoud Salhab, Marwan Elghitany, Shameed Sait, Syed Sibghat Ullah 외

Automatic speech recognition (ASR) is crucial for human-machine interaction in diverse applications like conversational agents, industrial robotics, call center automation, and automated subtitling. However, developing h…

Arabic Speech RecognitionAutomatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognition+2

Towards a methodology for evaluating automatic subtitling

2022-06-01 · EAMT 2022 6 · Alina Karakanta, Luisa Bentivogli, Mauro Cettolo, Matteo Negri 외

In response to the increasing interest towards automatic subtitling, this EAMT-funded project aimed at collecting subtitle post-editing data in a real use case scenario where professional subtitlers edit automatically ge…

Segmentation