Live data from Hacker News

NotebookLM's automatically generated podcasts are surprisingly effective

simonwillison.net

361–370 of 511 posts

Re: NotebookLM's automatically generated podcasts are surprisingly effective

#361
post #356

Is there a tool to do the opposite? I can't stand podcasts as a format (even if transcribed).

Google Gemini running in AI Studio accepts audio files, so you can upload a MP3 to it directly and prompt it to "rewrite this content as a casual blog post" (or whatever format you want) and it should work really well.

Or manually transcribe the podcast with Whisper (I use the MacWhisper app for this all the time) and then dump that transcript into an LLM and ask it to reformat that.

Re: NotebookLM's automatically generated podcasts are surprisingly effective

#362

Earlier quoted context omitted.

I keep seeing this asertion: "the robots will get there" (or its ilk), and it's starting to feel really weird to me. It's an article of faith -- we don't KNOW that they're going to get there. They're going to get better, almost certainly, but how much? How much gas is left in the tank for this technique? Honestly, I think the fact that every new "groundbreaking" news release about LLMs has come alongside a swath of d…

IMO it seems almost epistemologically impossible that LLM's following anything even resembling the current techniques will ever be able to comfortably out-perform humans at genuinely creative endeavours because they, almost by definition, cannot be "exceptional". If you think about how an LLM works, it's effectively going "given a certain input, what is the statistically average output that I should provide, given my…

Yes, LLMs are probably inherently limited, but the AI field in general is not necessarily limited, and possibly has the potential to be more genuinely creative than even most exceptional creative humans.

Re: NotebookLM's automatically generated podcasts are surprisingly effective

#363
post #85

Earlier quoted context omitted.

This is true but the quality frontier is not a single bar. For mainstream content the bar is high. For super-niche content, I wouldn’t be surprised if NotebookLM already competes with the existing pods. This will be the dynamic of generated art as it improves; the ease of use will benefit creators at the fringe. I bet we see a successful Harry Potter fanfic fully generated before we see a AAA Avengers movie or simila…

On the contrary, the mainstream eats any slop you put infront of it as long as it follows the correct form - one needs only look at cable news - the super niche content is that which requires deep thinking and novel insights. Or to put another way, I've heard much better ideas on a podcast made by undergrad CS students than on Lex Fridman.

Cable news viewership has been rapidly dwindling.

Re: NotebookLM's automatically generated podcasts are surprisingly effective

#364
post #283

Earlier quoted context omitted.

> The one and only point listening to a discussion about anything is that at least one of the speakers is someone who has an opinion that you may find interesting or refutable. No. Maybe that's true for you , but people enjoy learning in different ways, and some people learn best by listening to a discussion.

Assuming you are one of them, I’m curious about one thing (honest question, not meant to disrespect): does it not bother you at all to know that those voices do not belong to any human being? When I listen to a semi-adolescent girl’s voice explaining something with a lot of “like”s and an informal tone, the fact that I know this was AI-generated makes me feel disgusted in my stomach (I am serious, this is not suppose…

I'm not - I think I learn better by reading. But I know a lot of people who do prefer discussions, and I thought that the comment I replied to came off as arrogant and dismissive of the idea that anyone else might learn differently.

I've listened to a few NotebookLM samples but haven't used it myself, so I can't really speak to how creepy it is in practice. Probably pretty creepy! (I don't think that the female voice in the samples sounds "semi-adolescent," though, for what it's worth - both of the voices just sound like millennial podcasters to me.)

Re: NotebookLM's automatically generated podcasts are surprisingly effective

#365
post #283

Earlier quoted context omitted.

> The one and only point listening to a discussion about anything is that at least one of the speakers is someone who has an opinion that you may find interesting or refutable. No. Maybe that's true for you , but people enjoy learning in different ways, and some people learn best by listening to a discussion.

Unlikely. It's just that our brains are so fried by our smartphones/social media/24h of news/media consumption that we've lost the plot.

It's unlikely that some people prefer to learn by listening to discussions?

Re: NotebookLM's automatically generated podcasts are surprisingly effective

#366

Anyone making the argument that computers/LLMs can only create mediocre content, and can’t (or it will take a long time to) create content that humans will find exceptional, needs to go back and read the commentary re: chess bots and go bots over the past ten or twenty years. We went from “computers can’t beat humans” to “okay, computers can beat humans, but they play like computers” to “computers are coming up with…

Is there anything to the notion that in Go, success and failure are concrete, objective, and more or less easy to measure (or at least measured along the same kind of rules)? While it is computationally intractable to iterate through future moves to an end state, it’s still relatively easy to understand how well you’re doing at any point, and you measure that in basically the same way every game.

For some parts of language, that’s true: there’s grammar, there’s syntax, there’s patois, there’s argot—all these things seem accountable to words’ collective frequency within articulable groups of speakers, more-or-less-fully knowable on their own, and with success metrics that evolve but that do so through collective processes that models can measure and calibrate to. And indeed the models are great at those aspects of language.

“Succeeding” at writing is more than just “saying it well,” it’s also “having something worth saying” and “being worth listening to.” The second point is where things seem to get hazier for computable models. For sure there’s a set of facts that are more or less constant about the world, and well-reported. Science, repackaging history that’s already been done, lurid tales of crime—the stuff podcasts are made of! Not to mention the vast sea of data that sensor networks and automated research can produce—vast reservoirs of subtle truth that humans struggle to begin to mine for insight! It makes complete sense that this is computable stuff, and that computed writing might well be worth learning from.

But important writing—classically, anyway—seems to involve communicating new or idiosyncratic knowledge, and often reveals some of the process of developing it. The podcast Serial, for example, was a smash hit specifically because it didn’t rely on things that were part of the record—and because it reminded people how contingent memory and “truth” are. Bob Woodward writes things that are shamelessly tinted with Bob-Woodward-worldview, but people reveal important and true things only to Bob Woodward because they trust who he is and how he’s behaved for a lifetime (prominent longtime investigative journalist in the US, on the national security beat). Nassim Taleb seems to come up around here: in something like Antifragile his project wasn’t necessarily about new facts but about interpreting them in contrarian fashion and grouping those contrarian insights to synthesize a new theory.

Which brings us to the third component: “being worth listening to.” Writing is an act of communication: the writer matters. A parent hangs its child’s crayon drawing on the fridge not because it’s “authentic to the style of the kids’-crayon-drawing mode of visual art,” not because it’s novel or informative or even true-to-life, but because it came from a person they love. A “Dear John” letter devastates a soldier because it comes from a person with outsized part in their life and identity. Chinese publishers’ booths at trade shows are wall-to-wall translations of The Governance of China because it’s politically unwise not to. My favorite writers feel fresh: you feel elements of their personality come through. People have a special fetish for true crime—not that there’s any lack of fictitious crime to read about, but the fact that it happened to real humans potentiates the drama for these readers. It’s this aspect that I have a hard time understanding as computable (or commoditizable, I guess… are those similar phenomena?).

Already we seem to be drawing these distinctions in our collective reaction to LLM-stuff. We can’t wait to get hallucinations under control so we can chuck in gigantic boring contracts and internal wikis and financial reports, and get out comprehensible insight—but we roll our eyes at the tsunami of empty slop that’s overtaken Google results. We giggle at AI ventriloquism like this Neuro character [0], but die a little inside every time we read anodyne LLM-ish promotional copy and sameish AI art. First-level customer support seems like a perfect role for a chatbot—“turn it off and on again,” but nicely!—but people on the receiving end hate it [1] even for that task well-suited to it.

I’m only a layperson of course, but I wonder if any of those distinctions might be fruitful? Some of it I guess sums up to the old writing advice “show, don’t tell”—are there examples of machine writing showing promise in that way?

[0] https://m.youtube.com/@neurochron_fan_channel (video; brain rot)

[1] https://www.theregister.com/2024/07/09/gartner_simply_replac...

Re: NotebookLM's automatically generated podcasts are surprisingly effective

#367

NotebookLM's is incredibly good at generating the affect and structure of a quality podcast. This is in-line with all art, music, and video created by LMM at the moment. They are imitating a structure and affect, the quality of the content is largely irrelevant. I think the interesting thing is that most people don't really care, and AI is not to blame for that. Most books published today have the affect of a book, b…

> The reason so much writing, podcasting, and music is vulnerable to AI disruption is that quality has already become secondary. Commercial creative workers are vulnerable because there's a billions-of-dollars industry effort to copy their professional output and compete with them selling cheap knock-offs. I see this sort of convenient resignation all the time in the tech crowd... "creative workers only can blame the…

I'm confused about your point RE: infomercials et al - that's poor quality "content" that's been proliferating for, as you say, more than 50 years.

Is that not the work of commercial creative workers? Did it not exist pre AI? There's an argument to scale, certainly, but the idea that "things were better in the past before these > came out" is generally a suspect argument.

To your broader point - new tools for creating creative work come out all the time. Did we suffer greatly at the loss of image compositors when Photoshop arrived? On the flip side, did digital art gut painting and sculpture? Isn't this just another tool for creative expression?

Art is a way of seeing, not a way of creating. I don't think the technology is taking that away.

Re: NotebookLM's automatically generated podcasts are surprisingly effective

#368

Earlier quoted context omitted.

This is some insane catastrophizing. The value is that it turns it into a form factor that may be easier to consume, pay attention to, etc.

... insane catastrophizing." Nice unique phrase. Guessing you're not a LLM. ;^) The thing that is being offered is of no interested to me, as are almost any AI generated content. I'm a human, and am interested in what humans do and say and think. AI content offends my sensibilities at every level. I dismiss it without even thinking twice. So all those people who do podcast, music, art, whatever, with AI, well, you lo…

I will note this is slightly less an example of "AI generated" and more an example of "AI transformed". This takes existing, written by human documents or articles and transforms them into a podcast. Based on what you've written here, this shouldn't necessarily be in contradiction with your values, since you're still getting thoughts from other humans, and you can still pay money to the humans who made the original article, etc.

Re: NotebookLM's automatically generated podcasts are surprisingly effective

#369

NotebookLM's is incredibly good at generating the affect and structure of a quality podcast. This is in-line with all art, music, and video created by LMM at the moment. They are imitating a structure and affect, the quality of the content is largely irrelevant. I think the interesting thing is that most people don't really care, and AI is not to blame for that. Most books published today have the affect of a book, b…

> The reason so much writing, podcasting, and music is vulnerable to AI disruption is that quality has already become secondary. Commercial creative workers are vulnerable because there's a billions-of-dollars industry effort to copy their professional output and compete with them selling cheap knock-offs. I see this sort of convenient resignation all the time in the tech crowd... "creative workers only can blame the…

Have you never in your life enjoyed a pirated movie, game, book or music track?

Re: NotebookLM's automatically generated podcasts are surprisingly effective

#370

NotebookLM's is incredibly good at generating the affect and structure of a quality podcast. This is in-line with all art, music, and video created by LMM at the moment. They are imitating a structure and affect, the quality of the content is largely irrelevant. I think the interesting thing is that most people don't really care, and AI is not to blame for that. Most books published today have the affect of a book, b…

> The reason so much writing, podcasting, and music is vulnerable to AI disruption is that quality has already become secondary. I think that has always been the case, we just tend to compare today’s average stuff with the best stuff from earlier days. For example, most furniture pictures from the 60s and 70s are from upper middle class homes. If we listen music, we listen to Queen and not some local band from Alabam…

> I think that has always been the case, we just tend to compare today’s average stuff with the best stuff from earlier days.

I agree with this of course, because generally nobody remembers the bad stuff unless it was the worst. I beg to differ with music, though, because there's an opposing effect: we tend to be left with the most marketed music, which was usually a cheap knockoff of something interesting going on at the time. The shitty commercial knockoff becomes the "classic" while the people they were ripping off don't even get a wikipedia page.

Post reply on HN