Live data from Hacker News

Audapolis: Edit audio files by transcript, not waveform

github.com

71–80 of 87 posts

Re: Audapolis: Edit audio files by transcript, not waveform

#72

Call me a jerk, but anyone who is editing audio seriously, probably wants the waveform, no?

Podcasters are much less picky, with much more audio to process. For music or film, I would strongly agree.

It's also probably Good Enough for a first pass-through.

I'm stuck in editing hell right now, and it would be very nice to just visually scroll past a few pages of pre-episode bullshitting and be able to wipe out whole minutes at a stretch, without having to listen to the whole thing. Even at increased speed, it's a bit of a slog.

Re: Audapolis: Edit audio files by transcript, not waveform

#74
post #9

I've spent some of my free time over the past couple of months working on something similar. It's in a decent state but I need help from somebody who understands the .fcpxml format so you can export your edits to Davinci and FCP. Take a look at https://matcha.video

Looks useful. Does it export as a video file itself (e.g., mp4)? Thanks.

Re: Audapolis: Edit audio files by transcript, not waveform

#75

Call me a jerk, but anyone who is editing audio seriously, probably wants the waveform, no?

Podcasters are much less picky, with much more audio to process. For music or film, I would strongly agree.

as podcaster, yup. chucking 2hrs of audio in descript and removing 700 ums is golden

Re: Audapolis: Edit audio files by transcript, not waveform

#76

Earlier quoted context omitted.

Nah. It's just being a condescending dick.

To be fair, I didn't read it as condescending. This I understood it as a generic statement ender, like "idk" or "ymmv" or "i guess", i.e. something that you put at the end of a sentence when you don't know what more to say. Maybe it actually is generational?

Thanks that was my understanding too, but I hurt the boomers with "whatevs" lol

If it makes y'all feel better my gen alpha kid will hurt my feelings in different ways and get revenge for you =_=

Re: Audapolis: Edit audio files by transcript, not waveform

#77
post #10

Nice, are there plans to notarize the mac app? I built something similar here: https://bigwav.app

I just tried this out and it's very nice and easy to use. Thank you for sharing! I ended up copy-pasting the output from the messages page, which is 99% of the way to exporting a .txt file and my personal use case. Great work.

Re: Audapolis: Edit audio files by transcript, not waveform

#78
post #62

The other day I was using the voice memos app on iOS 18 and was surprised to find that it also supports editing the recording by transcript

I just upgrade to iOS 18 to try this and couldn't find it. How do you actually do it?

I think it generates the transcript automatically for new recordings but you can also edit a old one and then generate the transcript from there

Re: Audapolis: Edit audio files by transcript, not waveform

#79
post #77
post #10

Nice, are there plans to notarize the mac app? I built something similar here: https://bigwav.app

I just tried this out and it's very nice and easy to use. Thank you for sharing! I ended up copy-pasting the output from the messages page, which is 99% of the way to exporting a .txt file and my personal use case. Great work.

Thanks for the feedback! Maybe I should add a "download as .txt".

What do you typically do with the text on export? E.g. Do you parse the times?

Re: Audapolis: Edit audio files by transcript, not waveform

#80
post #79
post #77

Earlier quoted context omitted.

I just tried this out and it's very nice and easy to use. Thank you for sharing! I ended up copy-pasting the output from the messages page, which is 99% of the way to exporting a .txt file and my personal use case. Great work.

Thanks for the feedback! Maybe I should add a "download as .txt". What do you typically do with the text on export? E.g. Do you parse the times?

In my videography work I often do a separate audio-only interview to use as voice-over for the final video. I like to print out a transcript, mark the highlights, then go to the sound file and extract the snippets I liked. Extracting the snippets is a lot easier when I have timestamps printed out inline with the text at intervals of one or two minutes. In the case of bigWav, there were timestamps marked at only three or four points, so I had to go back and manually enter ten more marks to orient myself on the page. In addition, I used ChatGPT on an answer-by-answer basis to clean up the copy and add in punctuation for ease of reading. So there was an hour or two of data sanitizing needed to get everything ready to print out and use efficiently.
Post reply on HN