Live data from Hacker News

Transcribro: On-device Accurate Speech-to-text

github.com

61–66 of 66 posts

Re: Transcribro: On-device Accurate Speech-to-text

#61
post #3

Documentation severely lacking. I wanted to know whether this does streaming or only batch, as well as examples for integrating with Android apps.

It uses VAD and processes after it detects no speech for 3 seconds, so only batch. Examples for integrating with Android apps? Like apps that can use it? Pretty much any app that uses Android's SpeechRecognizer class if you set Transcribro as the user-selected speech recognizer or if the app uses Transcribro explicitly. For example, Google Maps uses the user-selected speech recognizer when it doesn't detect Google's speech services on the system.

Re: Transcribro: On-device Accurate Speech-to-text

#63
post #44

Earlier quoted context omitted.

You're right that it exists, but it's complete crap outside a quiet environment. Try to use it while walking around outside or in any semi-noisy area and it fails horribly (iPhone 13, so YMMV if you have a newer one). You cannot use an iPhone as a dictation device without reviewing the transcribed text, which IMO defeats the purpose of dictation. Meanwhile, i've gotten excellent results on the iPhone from a Whipser->…

I've never found real-time dictation software that doesn't need to be reviewed. I'm definitely waiting for Apple to upgrade their dictation software to the next generation -- I have my own annoyances with it -- but I haven't found anything else that works way better, in real time, on a phone, that runs in the background (like as part of the keyboard). You talk about Whisper but that doesn't even work in real time, mu…

What's the real-time requirement for? We may have different use cases, but it's not needed if I don't need to review the results. Speak -> Send, without reviewing the text, is the desired workflow. I.e. so you can compose messages without looking at your phone.

So yes, i'm not sure of alternate real-time solutions, but the non real-time solution of Whisper is much better for my real-world use case.

Re: Transcribro: On-device Accurate Speech-to-text

#65

Earlier quoted context omitted.

I looked in the GitHub issues and there's a closed issue for F-droid inclusion. The author states that F-droid "Doesn't meet their requirements" but doesn't elaborate. I wonder what F-droid is missing that they need so much?

Reason https://www.privacyguides.org/en/android/#f-droid

Author also points to https://privsec.dev/posts/android/f-droid-security-issues/

I'm not really sold on the argument... Also constant push/hype of GrapheneOS (and the "attitude" of it's devs) is mildly annoying...

Re: Transcribro: On-device Accurate Speech-to-text

#66

Earlier quoted context omitted.

Reason https://www.privacyguides.org/en/android/#f-droid

Author also points to https://privsec.dev/posts/android/f-droid-security-issues/ I'm not really sold on the argument... Also constant push/hype of GrapheneOS (and the "attitude" of it's devs) is mildly annoying...

[deleted]
Post reply on HN