paper-with-me

홈 › Papers

Earnings-21: A Practical Benchmark for ASR in the Wild

2021-04-22 · Miguel Del Rio, Natalie Delworth, Ryan Westerman, Michelle Huang, Nishchal Bhandari, Joseph Palakapilly, Quinten McNamara, Joshua Dong, Piotr Zelasko, Miguel Jette

Commonly used speech corpora inadequately challenge academic and commercial ASR systems. In particular, speech corpora lack metadata needed for detailed analysis and WER measurement. In response, we present Earnings-21, a 39-hour corpus of earnings calls containing entity-dense speech from nine different financial sectors. This corpus is intended to benchmark ASR systems in the wild with special attention towards named entity recognition. We benchmark four commercial ASR models, two internal models built with open-source tools, and an open-source LibriSpeech model and discuss their differences in performance on Earnings-21. Using our recently released fstalign tool, we provide a candid analysis of each model's recognition capabilities under different partitions. Our analysis finds that ASR accuracy for certain NER categories is poor, presenting a significant impediment to transcript comprehension and usage. Earnings-21 bridges academic and commercial ASR system evaluation and enables further research on entity modeling and WER on real world audio.

📄 PDF Abstract BibTeX arXiv:2104.11348

Code (2)

revdotcom/speech-datasets/tree/main/earnings21 공식 구현
revdotcom/fstalign

Tasks

named-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)NER

Similar Papers 제목 키워드 기반

Earnings-22: A Practical Benchmark for Accents in the Wild

2022-03-29 · Miguel Del Rio, Peter Ha, Quinten McNamara, Corey Miller 외

Modern automatic speech recognition (ASR) systems have achieved superhuman Word Error Rate (WER) on many common corpora despite lacking adequate performance on speech in the wild. Beyond that, there is a lack of real-wor…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Benchmarkingspeech-recognition+1

Contextual Earnings-22: A Speech Recognition Benchmark with Custom Vocabulary in the Wild

2026-03-28 · Berkin Durmus, Chen Cen, Eduardo Pacheco, Arda Okan 외 arxiv

The accuracy frontier of speech-to-text systems has plateaued on academic benchmarks.1 In contrast, industrial benchmarks and adoption in high-stakes domains suggest otherwise. We hypothesize that the primary difference …

Speech Recognition

Same Company, Same Signal: The Role of Identity in Earnings Call Transcripts

2024-12-23 · Ding Yu, Zhuo Liu, Hangfeng He

Post-earnings volatility prediction is critical for investors, with previous works often leveraging earnings call transcripts under the assumption that their rich semantics contribute significantly. To further investigat…

Attribute

Numerical Claim Detection in Finance: A New Financial Dataset, Weak-Supervision Model, and Market Analysis

2024-02-18 · Agam Shah, Arnav Hiray, Pratvi Shah, Arkaprabha Banerjee 외

In this paper, we investigate the influence of claims in analyst reports and earnings calls on financial market returns, considering them as significant quarterly events for publicly traded companies. To facilitate a com…

Trading through Earnings Seasons using Self-Supervised Contrastive Representation Learning

2024-09-25 · Zhengxin Joseph Ye, Bjoern Schuller

Earnings release is a key economic event in the financial markets and crucial for predicting stock movements. Earnings data gives a glimpse into how a company is doing financially and can hint at where its stock might go…

Algorithmic TradingRepresentation LearningSelf-Supervised Learning