Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

This cool. But are these audio files transcribed, or just provided?

I downloaded a couple here:

http://www.repository.voxforge1.org/downloads/SpeechCorpus/T...

and it didn't seem to have a log of words labeled each by timestamp offset into the audio recording - which is the vital part for training a recognizer. Am I missing something?



It's trickier than just matching word sounds, the sphinx docs are first rate: http://cmusphinx.sourceforge.net/wiki/tutorialam

It is very interesting but unfortunately just appears to be too much hassle for sane people to tackle (although it'd be extremely worthwhile if someone would innovate in this space and lower the barrier to entry - most are using commercial acoustic models with the FOSS software)




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: