paper-with-me

홈 › Papers

BanglaTalk: Towards Real-Time Speech Assistance for Bengali Regional Dialects

2025-10-07 · Jakir Hasan, Shubhashis Roy Dipta arxiv

Real-time speech assistants are becoming increasingly popular for ensuring improved accessibility to information. Bengali, being a low-resource language with a high regional dialectal diversity, has seen limited progress in developing such systems. Existing systems are not optimized for real-time use and focus only on standard Bengali. In this work, we present BanglaTalk, the first real-time speech assistance system for Bengali regional dialects. BanglaTalk follows the client-server architecture and uses the Real-time Transport Protocol (RTP) to ensure low-latency communication. To address dialectal variation, we introduce a dialect-aware ASR system, BRDialect, developed by fine-tuning the IndicWav2Vec model in ten Bengali regional dialects. It outperforms the baseline ASR models by 12.41-33.98% on the RegSpeech12 dataset. Furthermore, BanglaTalk can operate at a low bandwidth of 24 kbps while maintaining an average end-to-end delay of 4.9 seconds. Low bandwidth usage and minimal end-to-end delay make the system both cost-effective and interactive for real-time use cases, enabling inclusive and accessible speech technology for the diverse community of Bengali speakers. Code is available in https://github.com/Jak57/BanglaTalk

📄 PDF Abstract BibTeX arXiv:2510.06188

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Comprehending Real Numbers: Development of Bengali Real Number Speech Corpus

2018-03-27 · Md Mahadi Hasan Nahid, Md. Ashraful Islam, Bishwajit Purkaystha, Md. Saiful Islam

Speech recognition has received a less attention in Bengali literature due to the lack of a comprehensive dataset. In this paper, we describe the development process of the first comprehensive Bengali speech dataset on r…

speech-recognitionSpeech Recognition

OOD-Speech: A Large Bengali Speech Recognition Dataset for Out-of-Distribution Benchmarking

2023-05-15 · Fazle Rabbi Rakib, Souhardya Saha Dip, Samiul Alam, Nazia Tasnim 외

We present OOD-Speech, the first out-of-distribution (OOD) benchmarking dataset for Bengali automatic speech recognition (ASR). Being one of the most spoken languages globally, Bengali portrays large diversity in dialect…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)BenchmarkingDiversity+2

A Bengali HMM Based Speech Synthesis System

2014-06-16 · Sankar Mukherjee, Shyamal Kumar Das Mandal

The paper presents the capability of an HMM-based TTS system to produce Bengali speech. In this synthesis method, trajectories of speech parameters are generated from the trained Hidden Markov Models. A final speech wave…

Speech Synthesistext-to-speechText to Speech

BengaliSent140: A Large-Scale Bengali Binary Sentiment Dataset for Hate and Non-Hate Speech Classification

2026-01-27 · Akif Islam, Sujan Kumar Roy, Md. Ekramul Hamid arxiv

Sentiment analysis for the Bengali language has attracted increasing research interest in recent years. However, progress remains constrained by the scarcity of large-scale and diverse annotated datasets. Although severa…

Sentiment Analysis

Bengali Common Voice Speech Dataset for Automatic Speech Recognition

2022-06-28 · Samiul Alam, Asif Sushmit, Zaowad Abdullah, Shahrin Nakkhatra 외

Bengali is one of the most spoken languages in the world with over 300 million speakers globally. Despite its popularity, research into the development of Bengali speech recognition systems is hindered due to the lack of…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)DiversitySentence+2