Creating an Aligned Corpus of Sound and Text: The Multimodal Corpus of Shakespeare and Milton
In this work we present a corpus of poems by William Shakespeare and John Milton that have been enriched with readings from the public domain. We have aligned all the lines with their respective audio segments, at the line, word, syllable and phone level, and we have included their scansion. We make a basic visualization platform for these poems and we conclude by conjecturing possible future directions.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
Sound-aligned corpus of Udmurt dialectal texts
SPEECH-COCO: 600k Visually Grounded Spoken Captions Aligned to MSCOCO Data Set
This paper presents an augmentation of MSCOCO dataset where speech is added to image and text. Speech captions are generated using text-to-speech (TTS) synthesis resulting in 616,767 spoken captions (more than 600h) pair…
text-to-speechText to SpeechMultiNews: A Web collection of an Aligned Multimodal and Multilingual Corpus
Integrating Natural Language Processing (NLP) and computer vision is a promising effort. However, the applicability of these methods directly depends on the availability of a specific multimodal data that includes images…
ArticlesContent-Based Image RetrievalImage RetrievalMachine Translation+1A Recipe for Creating Multimodal Aligned Datasets for Sequential Tasks
Many high-level procedural tasks can be decomposed into sequences of instructions that vary in their order and choice of tools. In the cooking domain, the web offers many partially-overlapping text and video recipes (i.e…
DescriptiveSynthScribe: Deep Multimodal Tools for Synthesizer Sound Retrieval and Exploration
Synthesizers are powerful tools that allow musicians to create dynamic and original sounds. Existing commercial interfaces for synthesizers typically require musicians to interact with complex low-level parameters or to …
Multimodal Deep LearningRetrieval