As an aside, it seems you're interested in speech recognition, or speech to text, not voice recognition. Voice recognition is a different problem, where the particular speaker needs to be recognized from voice.
It’s pretty hard to blame lay people when speech reco products like Dragon are widely marketed as “voice recognition”.
If you want to be clear and not just pedantic just call it speaker recognition.