Live data from Hacker News

Audapolis: Edit audio files by transcript, not waveform

github.com

61–70 of 87 posts

Re: Audapolis: Edit audio files by transcript, not waveform

#61
post #13

Earlier quoted context omitted.

I love Descript. Their "convert to studio quality" feature is better than Adobe's and ElevenLabs, in my experience. I wondered if this particular feature was really worth paying for so I was happy that I found Audapolis.

What does that feature do?

What I particularly like about the Descript version (though it is overdone as mentioned) is that it reduces or eliminates the pesky S sounds and P sounds (called sibilances and plosives) that you get when you talk into a microphone and you're not perfectly distanced from it.

I haven't found another app that reduces or removes these.

Re: Audapolis: Edit audio files by transcript, not waveform

#63
post #59

Combine this with the tech to generate new audio matching the speaker's voice profile, and you've really got something cool.

It's difficult to do this for a video (and probably wouldn't look that nice)

Descript does an okay job of this too.

Re: Audapolis: Edit audio files by transcript, not waveform

#64

Earlier quoted context omitted.

and there's still no commercial product for synthesizing video to sync lip movements to edited transcript like all the scary proof of concepts that turned the president into a puppet Maybe there's not much value in editing what someone said after all

commercial value: no criminal value... maybe? There are absolutely scams right now that use deepfakes to trick people.

> commercial value: no

Of course there is commercial value. The cost of reshooting video materials is huge. You made an advert mentioning 3 features, but by the time the product is about to be released one of them got dropped or even worse changed? Congrats, you need to get the talent and the studio rebooked, you need to find a new tech crew, who need to set up again. Probably things won't cut seamlessly so you need to re-record the whole thing.

Re: Audapolis: Edit audio files by transcript, not waveform

#66

Earlier quoted context omitted.

You're being insecure. It's not rude to disagree. Also, there's often no perfect combo of words, there's a spectrum of options and you just pick an operating point. Transcription is a longer word than "word" so there's a tradeoff. It doesn't feel like a chasm to me.

in some context "whatever" can be it evens out, but in others it can be "your opinion doesn't matter". At any rate when I first read it I thought it was going to be some sort LLM thing where you said "remove the third bridge and increase pitch by one octave in the outro" and it would give you back an edited mp4 which you could then listen and cringe to and sometimes say "whoa, that's amazing!"

lol, great description of what I thought too. Someone's going to do it...

Re: Audapolis: Edit audio files by transcript, not waveform

#67
post #64

Earlier quoted context omitted.

commercial value: no criminal value... maybe? There are absolutely scams right now that use deepfakes to trick people.

> commercial value: no Of course there is commercial value. The cost of reshooting video materials is huge. You made an advert mentioning 3 features, but by the time the product is about to be released one of them got dropped or even worse changed? Congrats, you need to get the talent and the studio rebooked, you need to find a new tech crew, who need to set up again. Probably things won't cut seamlessly so you need…

Potentially also for syncing lip movement for content dubbed in a foreign language. For me when watching foreign media dubbed to English the discrepancy is very noticeable and quite distracting, no matter how well the dub is written, performed, and edited to match the timing of the original.

Re: Audapolis: Edit audio files by transcript, not waveform

#68
post #59

Combine this with the tech to generate new audio matching the speaker's voice profile, and you've really got something cool.

It's difficult to do this for a video (and probably wouldn't look that nice)

This is for just audio, though?

Re: Audapolis: Edit audio files by transcript, not waveform

#70

I remember when Adobe demoed this idea of being able to edit waveforms by the recognized text back in 2016 and it was pretty mind blowing for the time. https://youtu.be/I3l4XLZ59iw EDIT: I could also definitely see Audapolis being useful if you could integrate it into a podcast's post processing flow (volume normalization, de-essing) by recognizing certain verbal tics and automatically removing them from the audio su…

This workflow is exactly what Descript does. Transcript-based editing, filler word removal, noise reduction, volume normalization, Overdub spoken word correction using the speaker’s voice, eye gaze correction for video, etc.

Disclaimer: I work at Descript

Post reply on HN