I am always skeptical of benchmarks that show perfect scores, especially when they come from the company selling the product. It feels like everyone claims to have solved conversational timing these days. I guess we will see if it is actually any good.
Different industry, but our marketing guy once said "You know what this [perfect] metric means? We can never use it in marketing because it's not believable"
Show HN: Sparrow-1 – Audio-native model for human-level turn-taking without ASR
11–20 of 50 posts
Re: Show HN: Sparrow-1 – Audio-native model for human-level turn-taking without ASR
#12Common ...
Re: Show HN: Sparrow-1 – Audio-native model for human-level turn-taking without ASR
#13Re: Show HN: Sparrow-1 – Audio-native model for human-level turn-taking without ASR
#14Literally no way to sign up to try. Put my email and password and it puts me into some wait list despite the video saying I could try the model today. That's what makes me mad about these kind of releases is that the marketing and the product don't talk together.
Re: Show HN: Sparrow-1 – Audio-native model for human-level turn-taking without ASR
#15Metric | Sparrow-1 Precision 100% Recall 100% Common ...
Re: Show HN: Sparrow-1 – Audio-native model for human-level turn-taking without ASR
#16Metric | Sparrow-1 Precision 100% Recall 100% Common ...
If you watch the demo video you can see how they would get this: the model is not aggressive enough. While it doesn't cut you off, which is nice, it also always waits an uncanny amount of time to chime in.
Re: Show HN: Sparrow-1 – Audio-native model for human-level turn-taking without ASR
#17I tried talking to Claude today. What a nightmare. It constantly interrupts you. I don’t mind if Claude wants to spend ten seconds thinking about its reply, but at least let ME finish my thought. Without decent turn-taking, the AI seems impolite and it’s just an icky experience. I hope tech like this gets widely distributed soon because there are so many situations in which I would love to talk with a model. If only…
Re: Show HN: Sparrow-1 – Audio-native model for human-level turn-taking without ASR
#18Re: Show HN: Sparrow-1 – Audio-native model for human-level turn-taking without ASR
#19Earlier quoted context omitted.
Different industry, but our marketing guy once said "You know what this [perfect] metric means? We can never use it in marketing because it's not believable"
Just include some noise, it’s like the most available resource in the universe