paper-with-me

Papers

100,000 Podcasts: A Spoken English Document Corpus

2020-12-01 · COLING 2020 8 · Ann Clifton, Sravana Reddy, Yongze Yu, Aasish Pappu, Rezvaneh Rezapour, Hamed Bonab, Maria Eskevich, Gareth Jones, Jussi Karlgren, Ben Carterette, Rosie Jones

Podcasts are a large and growing repository of spoken audio. As an audio format, podcasts are more varied in style and production type than broadcast news, contain more genres than typically studied in video data, and are more varied in style and format than previous corpora of conversations. When transcribed with automatic speech recognition they represent a noisy but fascinating collection of documents which can be studied through the lens of natural language processing, information retrieval, and linguistics. Paired with the audio files, they are also a resource for speech processing and the study of paralinguistic, sociolinguistic, and acoustic aspects of the domain. We introduce the Spotify Podcast Dataset, a new corpus of 100,000 podcasts. We demonstrate the complexity of the domain with a case study of two tasks: (1) passage search and (2) summarization. This is orders of magnitude larger than previous speech corpora used for search and summarization. Our results show that the size and variability of this corpus opens up new avenues for research.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

3D Facial Landmark LocalizationAutomatic Speech RecognitionAutomatic Speech Recognition (ASR)Facial Expression Recognition (FER)Highlight DetectionInformation RetrievalRetrievalspeech-recognitionSpeech Recognition

Methods 이 논문이 사용한 방법론

Où trouver le numéro de voyageur fréquent Emirates ? | Numéro client Vous avez un compte Skywards Emirates et souhaitez récupérer votre identifiant fidélité ? Le moyen le plus simple est de contacter l’assistance Emirates au +33 1 59 00 29 47. Ce…
[Contact~us] How do I really get through Qatar Airways? To connect with a live agent at Qatar Airways,☎️+1-801-(855)-(5905) or +1-804-(853)-(9001)✅ you can call their customer service line at ☎️+1-801-(855)-(5905) or…
How can I speak to a Lufthansa representative fast? Over 120 million people fly with Lufthansa Airlines each year, and many of them need real-time assistance with booking, cancellations, seat upgrades, or other travel issues.…
Ways to Speak Coinbase Wallet Support Number: Phone, Email, and Chat Support Options 설명 없음
QB Expert Guide: How Do I Contact Intuit QuickBooks Online Support 설명 없음
Teléfono®:¿Cómo llamar a Copa desde Costa Rica? 설명 없음
CoinSpot app not working — how to fix it? 설명 없음
{{Chiama Lufthanssa}} Come telefonare a Lufthansa? 설명 없음

Similar Papers 제목 키워드 기반

Cem Mil Podcasts: A Spoken Portuguese Document Corpus For Multi-modal, Multi-lingual and Multi-Dialect Information Access Research

2022-09-23 · Ekaterina Garmash, Edgar Tanaka, Ann Clifton, Joana Correia 외

In this paper we describe the Portuguese-language podcast dataset we have released for academic research purposes. We give an overview of how the data was sampled, descriptive statistics over the collection, as well as i…

DescriptiveGenre classification

It's What You Say and How You Say It: Exploring Textual and Audio Features for Podcast Data

2022-01-16 · ACL ARR January 2022 1 · Anonymous

Podcasts are relatively new media in the form of spoken documents or conversations with a wide range of topics, genres, and styles. With a massive increase in the number of podcasts and their listener base, it is benefic…

TAG

Current Challenges and Future Directions in Podcast Information Access

2021-06-17 · Rosie Jones, Hamed Zamani, Markus Schedl, Ching-Wei Chen 외

Podcasts are spoken documents across a wide-range of genres and styles, with growing listenership across the world, and a rapidly lowering barrier to entry for both listeners and creators. The great strides in search and…

Merkel Podcast Corpus: A Multimodal Dataset Compiled from 16 Years of Angela Merkel's Weekly Video Podcasts

2022-05-24 · Debjoy Saha, Shravan Nayak, Timo Baumann

We introduce the Merkel Podcast Corpus, an audio-visual-text corpus in German collected from 16 years of (almost) weekly Internet podcasts of former German chancellor Angela Merkel. To the best of our knowledge, this is …

Face DetectionFace GenerationSpeaker RecognitionTalking Face Generation

Merkel Podcast Corpus: A Multimodal Dataset Compiled from 16 Years of Angela Merkel’s Weekly Video Podcasts

2022-06-01 · LREC 2022 6 · Debjoy Saha, Shravan Nayak, Timo Baumann

We introduce the Merkel Podcast Corpus, an audio-visual-text corpus in German collected from 16 years of (almost) weekly Internet podcasts of former German chancellor Angela Merkel. To the best of our knowledge, this is …

Face DetectionFace GenerationSpeaker RecognitionTalking Face Generation