Speech recognition as a technology, has always appeared to move slowly although with the advent of mobile popularity, the technology is becoming increasingly popular.
Is anyone doing anything like this?
1–10 of 40 posts
Speech recognition as a technology, has always appeared to move slowly although with the advent of mobile popularity, the technology is becoming increasingly popular.
Is anyone doing anything like this?
There are two different but correlated fields: Speech recognition and Natural Language Understanding.Speech Recognition is easier if the scope is minimised, that is, if th system knows which subset of keywords of orders to recognized. But recognizing an open scope, including different accents, slang, etc is a much more difficult task.
It may also be possible to automate the entire process as we have both the audio and the words spoken at a particular time.
Take it a step further, we have millions of sung songs with lyrics that can also be used. Its a gold mine of information that can be repurposed.
Well I am sure they would do, though subtitles aren't the most reliable source for movie dialog. Often the dialog is altered subtly to fit the space and timing requirements.
Well I am sure they would do, though subtitles aren't the most reliable source for movie dialog. Often the dialog is altered subtly to fit the space and timing requirements.
How about music lyrics?
I wasn't arguing against movies, just that subtitles rather than a final script isn't the best data source.
For the purposes of speech recognition, songs strike me as being particularly noisy .