File size: 365 Bytes
ea4fd8d |
1 2 3 4 5 6 |
Personal speech to text model
-----------------------------
Speech to Text models often do not understand my accent, so I fine tuned this one from "openai/whisper-medium.en" using about 1000 recordings of my voice, comprising of about 2h of recordings. The system goes from ~9% WER to ~5% WER.
Do not download unless you have exactly my accent (North-East Italy). |