Live data from Hacker News

Show HN: We are 2.4x faster than Open AI latest SST model

news.ycombinator.com

1–6 of 6 posts

Show HN: We are 2.4x faster than Open AI latest SST model

#1
Speech-to-Text technology is crucial for audio processing applications, but most solutions are limited by speed, accuracy, or feature constraints. What if you could transcribe audio with higher accuracy, better features, and significantly faster performance?

We benchmarked OpenAI's newest gpt-4o-transcribe model against JigsawStack STT across real-world scenarios, including performance, accuracy, feature set, and multilingual capabilities. Here's what we found:

Performance:

- JigsawStack processes audio ~2.4x faster than OpenAI's model across all audio lengths

- This speed advantage remains consistent with both short clips and longer audio files

File Support:

- OpenAI limits files to 25MB and 25 minutes of audio per request

- JigsawStack handles files up to 100MB and 4 hours of audio per request

Advanced Features:

- Speaker Recognition: JigsawStack accurately identifies up to 50 different speakers with timestamps

- Detailed Timestamps: Provides sentence-level timing data by default

- Translation: Built-in support for translating audio into 100+ languages

Real-world Accuracy:

- Achieved 100% accuracy on noisy audio samples where OpenAI scored 94%

- Properly identified repeated phrases and sentence boundaries that OpenAI missed

Our team has focused on building a specialized solution that outperforms in nearly every metric that matters for real-world applications.

Check out the full breakdown here: http://jigsawstack.com/blog/openai-audio-stt-vs-jigsawstack-...

Let us know what you think! We'd love feedback from the HN community - particularly from those working with audio transcription at scale.

Re: Show HN: We are 2.4x faster than Open AI latest SST model

#4
post #2

Show HN is for things you've made other people can try and it excludes most reading material, take a look at https://news.ycombinator.com/showhn.html You can post benchmarks, blogposts, etc without the Show HN prefix though.

I think you can try the model

Re: Show HN: We are 2.4x faster than Open AI latest SST model

#5
post #2

Show HN is for things you've made other people can try and it excludes most reading material, take a look at https://news.ycombinator.com/showhn.html You can post benchmarks, blogposts, etc without the Show HN prefix though.

We encourage everyone to try out our model. Here is a link to get started: https://jigsawstack.com/speech-to-text

Re: Show HN: We are 2.4x faster than Open AI latest SST model

#6
post #5
post #2

Show HN is for things you've made other people can try and it excludes most reading material, take a look at https://news.ycombinator.com/showhn.html You can post benchmarks, blogposts, etc without the Show HN prefix though.

We encourage everyone to try out our model. Here is a link to get started: https://jigsawstack.com/speech-to-text

That's great but the thing you're linking and describing is still a blog post/benchmark and you also just had an actual Show HN https://news.ycombinator.com/item?id=43368327

You should still take a look at those Show HN rules and also https://news.ycombinator.com/item?id=22336638, the spirit of the thing is showing cool stuff you've built and when it starts sounding like obvious promotion it tends to work less well. It's in your own interest to categorize/frame your own postings better - they'll get better discussion that way.