Hindi Marathi Code-Switched Speech Recognition
preprint
OA: closed
CC-BY-4.0
Abstract
For Automatic Speech Recognition (ASR) and Natural Language Processing (NLP) systems, code-switching—the habit of alternately speaking in several languages within a single conversation—offers special difficulties and possibilities. ASR systems have to efficiently manage language transitions as multilingual communication gets more common if we want real-time speech recognition. This work investigates innovative approaches for processing code-switched audio, solves the dearth of multilingual datasets, and assesses several technologies applied to identify and analyze mixed-language speech. Emphasizing Hindi-Marathi code-switching, we present a dynamic language-switching architecture leveraging reinforcement learning methods including Q-Learning and Deep Q-Networks (DQN) to improve language transition identification. Moreover, we present a dataset especially meant for multilingual voice recognition and evaluate ASR performance with Character Error Rate (CER) and Word Error Rate (WER). Our study reveals current constraints and provides future directions to improve ASR adaptation, therefore guaranteeing more accurate and strong recognition in many multilingual settings.
My notes (saved in your browser only)
Citation neighborhood (no data yet)
We don't have any in-corpus citations linked to this paper yet. This is a recent paper (2025) — citers typically take a year or two to land, and the OpenAlex reference graph may still be filling in.
Source provenance
- europepmc
- last seen: 2026-05-20T01:45:00.602351+00:00
- unpaywall
- last seen: 2026-05-28T02:00:01.590549+00:00
License: CC-BY-4.0