An opensource speech-to-text software written in tensorflow
-
Updated
Oct 15, 2022 - Python
An opensource speech-to-text software written in tensorflow
Created an ASR (Automatic Speech Recognition) system that takes in individual recordings. Each recording represents a sentence composed of 5-10 English language digits, separated by adequate pauses. The system involves segmenting the sentence using a classifier, differentiating between background and foreground sounds.
Human-verified label-noise annotations for the Mobvoi Hotwords corpus (OpenSLR SLR87). Full listening audit of all 21,282 test positives, dev-set flagging, and the audit scripts. Companion to an ICASSP 2027 submission.
To associate your repository with the openslr topic, visit your repo's landing page and select "manage topics."