So far, we tried Google's Speech-to-text and Azure's speech to text, but both are less accurate and struggle with custom phrases.
We're working with large audio files (>30 minutes, >25 MB), but we can change how we parse up the files and the file types. All of our files are in Google Cloud Storage right now.
Curious if anyone has recommendations?