Live data from Hacker News

Audapolis: Edit audio files by transcript, not waveform

github.com

51–60 of 87 posts

Re: Audapolis: Edit audio files by transcript, not waveform

#51

Earlier quoted context omitted.

Nah. It's just being a condescending dick.

To be fair, I didn't read it as condescending. This I understood it as a generic statement ender, like "idk" or "ymmv" or "i guess", i.e. something that you put at the end of a sentence when you don't know what more to say. Maybe it actually is generational?

If you really want to be that charitable.

Re: Audapolis: Edit audio files by transcript, not waveform

#52

Earlier quoted context omitted.

I remember people saying at the time that “this is the point at which voice recordings can not be trusted any longer”. And then, like you said nothing happened kind of for a few years until the current AI/ML tech got to where it is currently at.

and there's still no commercial product for synthesizing video to sync lip movements to edited transcript like all the scary proof of concepts that turned the president into a puppet Maybe there's not much value in editing what someone said after all

HeyGen allows you to do for this in a few ways

Re: Audapolis: Edit audio files by transcript, not waveform

#54

And here I was expecting that I could edit the text and the app would change the audio file to say what I had typed...

Can I ask what this tool does? I was trying to figure it out (the GitHub page isn't terribly clear) and came to the same conclusion you did (delete a chunk of the transcript and the tool would delete that audio). I think I just lack experience in this area. I've used Audacity to cut out parts of audio / splice together two clips and that's about it, so I clearly don't have enough background to understand what this to…

It does exactly what you think it does. You can cut parts of the original file without having to edit the waveform (like you would in Audacity). Instead, select the parts directly just like you would in a text editor.

What it does not do is generate new words (ie you type a sentence and it adds that to your file as voice).

Re: Audapolis: Edit audio files by transcript, not waveform

#55
post #13

Earlier quoted context omitted.

I love Descript. Their "convert to studio quality" feature is better than Adobe's and ElevenLabs, in my experience. I wondered if this particular feature was really worth paying for so I was happy that I found Audapolis.

What does that feature do?

Its machine learning powered noise reduction + compressor + eq + normalize combo effect. Works ok. Results in quite a bit overdone “studio” sound. I think trend in mixing is leaning much more natural (less tweaked) nowdays. But for no work it might be impressive. Probably works in internal corpo presentations well.

Re: Audapolis: Edit audio files by transcript, not waveform

#56

Earlier quoted context omitted.

I remember people saying at the time that “this is the point at which voice recordings can not be trusted any longer”. And then, like you said nothing happened kind of for a few years until the current AI/ML tech got to where it is currently at.

and there's still no commercial product for synthesizing video to sync lip movements to edited transcript like all the scary proof of concepts that turned the president into a puppet Maybe there's not much value in editing what someone said after all

This is used pretty often, you probably just don't notice it.

Re: Audapolis: Edit audio files by transcript, not waveform

#57

Earlier quoted context omitted.

If you're trying to get attention, copy should be clear to all readers. The fact that you did not misread it in no way demonstrates that others won't. And why the rude response?

You're being insecure. It's not rude to disagree. Also, there's often no perfect combo of words, there's a spectrum of options and you just pick an operating point. Transcription is a longer word than "word" so there's a tradeoff. It doesn't feel like a chasm to me.

in some context "whatever" can be it evens out, but in others it can be "your opinion doesn't matter".

At any rate when I first read it I thought it was going to be some sort LLM thing where you said "remove the third bridge and increase pitch by one octave in the outro" and it would give you back an edited mp4 which you could then listen and cringe to and sometimes say "whoa, that's amazing!"

Post reply on HN