Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

Would it be difficult to make a program that sits in the audio stack (probably on Windows) and tries to run voice recognition on everything coming over the line and then renders anything that sounds like speech at the bottom of the screen?


Yes. Youtube tries (tried?) to autocaption videos, and the results were infamously bad.


Yes. Voice recognition for arbitrary voices, accents, and lexicons is not as reliable as you imagine, much less in real-time.




Consider applying for YC's Fall 2026 batch! Applications are open till July 27.

Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: