Your Adaptive Explainable AI Pronunciation Training Companion
-
Updated
Aug 17, 2026 - Python
Your Adaptive Explainable AI Pronunciation Training Companion
Deployed Facebook's wav2vec2-large-90h model to transcribe 4,076 .mp3 files
This repo provides step by step process from sctatch to fine tune facebook's wav2vec2-large model using transformers
An ASR System that can transcribe Bengali speech with regional dialects
A fine-tuned Wav2Vec2-based Automatic Speech Recognition (ASR) system with data augmentation, efficient training, and transcription capabilities. Supports local and Mozilla Common Voice datasets, with evaluation via Word Error Rate (WER). 🚀
Add a description, image, and links to the wav2vec2-large-960h topic page so that developers can more easily learn about it.
To associate your repository with the wav2vec2-large-960h topic, visit your repo's landing page and select "manage topics."