We've been using this at work to get inside of our customer's perspective. It's helpful to throw eg a bunch of point-of-sale data sync challenges into Notebook LM and eg pass a 10 minute audio to the team so they can understand where our work fits in.
Pushing the frontiers of audio generation
81–90 of 114 posts
Re: Pushing the frontiers of audio generation
#82Earlier quoted context omitted.
"Surely my genAI product won't be used to spam zero-effort slop all over the internet!" - guy whose genAI product will definitely be used to spam zero-effort slop all over the internet.
I think their main target is corporate creative jobs. Background music to ads/videos/etc. And just like with all AI, they will eat the jobs that support the rest of the system, making it a one and done. It will give a one time boost, and then be stuck at that level because creatives won't have the jobs that allowed them to add to the domain. In this case new music styles. New techniques. It's literally eating the see…
Re: Pushing the frontiers of audio generation
#83While it is impressive and I like to follow the advancements in this field, it is incredibly frustrating to listen to. I can't put my finger on why exactly. It's definitely closer to human-sounding, but the uncanny valley is so deep here that I find myself thinking "I just want the point, not the fake personality that is coming with it". I can't make it through a 30s demo.
This is good, but certainly not yet great.
Re: Pushing the frontiers of audio generation
#84While it is impressive and I like to follow the advancements in this field, it is incredibly frustrating to listen to. I can't put my finger on why exactly. It's definitely closer to human-sounding, but the uncanny valley is so deep here that I find myself thinking "I just want the point, not the fake personality that is coming with it". I can't make it through a 30s demo.
Re: Pushing the frontiers of audio generation
#85While it is impressive and I like to follow the advancements in this field, it is incredibly frustrating to listen to. I can't put my finger on why exactly. It's definitely closer to human-sounding, but the uncanny valley is so deep here that I find myself thinking "I just want the point, not the fake personality that is coming with it". I can't make it through a 30s demo.
In general people find the back and forth between the "hosts" engaging and also gives people time to digest the contents.
Re: Pushing the frontiers of audio generation
#86While it is impressive and I like to follow the advancements in this field, it is incredibly frustrating to listen to. I can't put my finger on why exactly. It's definitely closer to human-sounding, but the uncanny valley is so deep here that I find myself thinking "I just want the point, not the fake personality that is coming with it". I can't make it through a 30s demo.
Re: Pushing the frontiers of audio generation
#87Earlier quoted context omitted.
Agreed. To me it sounds like bad voice-over actors reading from a script. So the natural parts of a conversation where you might say the wrong thing and step back to correct yourself are all gone. Impressive for sure.
every step of technological advancement builds on top of the previous one. now it's bad voice actors, in 2 years it'll be great ones
Re: Pushing the frontiers of audio generation
#88While it is impressive and I like to follow the advancements in this field, it is incredibly frustrating to listen to. I can't put my finger on why exactly. It's definitely closer to human-sounding, but the uncanny valley is so deep here that I find myself thinking "I just want the point, not the fake personality that is coming with it". I can't make it through a 30s demo.
In similar vein, I’m glad they told me it was a funny story, because otherwise I wouldn’t have known.
Re: Pushing the frontiers of audio generation
#89Earlier quoted context omitted.
We're used to hearing some kind of identity behind voices -- we unconsciously sense clusters of vocabulary, intonation patterns, ticks, frequent interruption vs quiet patience, silence tolerance, response patterns to various triggers, etc that communicate a coherent person of some kind. We may not know that a given speaker is a GenX Methodist from Wisconsin that grew up at skate parks in the suburbs, but we hear clus…
Whether this stops at the uncanny valley or progresses to specific "AI celebrity" voices, I'm left thinking the engineers involved in this never stopped to think carefully about whether this ought to be done in the first place.
Re: Pushing the frontiers of audio generation
#90Earlier quoted context omitted.
Whether this stops at the uncanny valley or progresses to specific "AI celebrity" voices, I'm left thinking the engineers involved in this never stopped to think carefully about whether this ought to be done in the first place.
"Surely my genAI product won't be used to spam zero-effort slop all over the internet!" - guy whose genAI product will definitely be used to spam zero-effort slop all over the internet.