Live data from Hacker News

Google AI Overviews cite YouTube more than any medical site for health queries

theguardian.com

191–200 of 214 posts

Re: Google AI Overviews cite YouTube more than any medical site for health queries

#191

Earlier quoted context omitted.

I think we hit peak AI improvement velocity sometime mid last year. The reality is all progress was made using a huge backlog of public data. There will never be 20+ years of authentic data dumped on the web again. I've hoped against but suspected that as time goes on LLMs will become increasingly poisoned by the the well of the closed loop. I don't think most companies can resist the allure of more free data as bitt…

> I don't think most companies can resist the allure of more free data as bitter as it may taste. Mercor, Surge, Scale, and other data labelling firms have shown that's not true. Paid data for LLM training is in higher demand than ever for this exact reason: Model creators want to improve their models, and free data no longer cuts it.

I did read or listen on a podcast about the booming business of AI data sets late last year. I'm sure you are right.

Doesn't change my point, I still don't think they can resist pulling from the "free" data. Corps are just too greedy and next quarter focused.

Re: Google AI Overviews cite YouTube more than any medical site for health queries

#192
post #46
post #14

Heavy Gemini user here, another observation: Gemini cites lots of "AI generated" videos as its primary source, which creates a closed loop and has the potential to debase shared reality. A few days ago, I asked it some questions on Russia's industrial base and military hardware manufacturing capability, and it wrote a very convincing response, except the video embedded at the end of the response was an AI generated o…

All of that and you're still a heavy user? Why would google change how Gemini works if you keep using it despite those issues?

Every single LLM out there suffers from this.

Re: Google AI Overviews cite YouTube more than any medical site for health queries

#193
post #14

Heavy Gemini user here, another observation: Gemini cites lots of "AI generated" videos as its primary source, which creates a closed loop and has the potential to debase shared reality. A few days ago, I asked it some questions on Russia's industrial base and military hardware manufacturing capability, and it wrote a very convincing response, except the video embedded at the end of the response was an AI generated o…

Try Kagi’s Research agent if you get a chance. It seems to have been given the instruction to tunnel through to primary sources, something you can see it do on reasoning iterations, often in ways that force a modification of its working hypothesis.

I suspect Kagi is running a multi-step agentic loop there, maybe something like a LangGraph implementation that iterates on the context. That burns a lot of inference tokens and adds latency, which works for a paid subscription but probably destroys the unit economics for Google's free tier. They are likely restricted to single-pass RAG at that scale.

Re: Google AI Overviews cite YouTube more than any medical site for health queries

#194

Earlier quoted context omitted.

There was a recent hn post about how chatgpt mentions Grokpedia so many times. Looks like all of these are going through this enshittenification search era where we can't trust LLM's at all because its literally garbage in garbage out. Someone had mentioned Kagi assistant in here and although they use API themselves but I feel like they might be able to provide their custom search in between, so if anyone's from Kagi…

Correct, Kagi Assistant uses Kagi Search - with all modifications user made (eg blocked domains, lenses etc).

> with all modifications user made

I've been wondering about that! Nice to have confirmation

Re: Google AI Overviews cite YouTube more than any medical site for health queries

#195

Earlier quoted context omitted.

[flagged]

"Doing your own research" is back on the menu boys!

I'll insist the surgeon follows ChatGPTs plan for my operation next time I'm in theatre.

By the end of the year AI will be actually doing the surgery, when you look at the recent advancements in robotic hands, right bros?

Re: Google AI Overviews cite YouTube more than any medical site for health queries

#196

Earlier quoted context omitted.

Try Kagi’s Research agent if you get a chance. It seems to have been given the instruction to tunnel through to primary sources, something you can see it do on reasoning iterations, often in ways that force a modification of its working hypothesis.

I suspect Kagi is running a multi-step agentic loop there, maybe something like a LangGraph implementation that iterates on the context. That burns a lot of inference tokens and adds latency, which works for a paid subscription but probably destroys the unit economics for Google's free tier. They are likely restricted to single-pass RAG at that scale.

> works for a paid subscription but probably destroys the unit economics for Google's free tier

Anyone relying on Google's free tier to attempt any research is getting what they pay for.

Re: Google AI Overviews cite YouTube more than any medical site for health queries

#197

Earlier quoted context omitted.

> YouTube channels with AI generated videos have exploded in sheer quantity, and I think majority of the new channels and videos uploaded to YouTube might actually be AI; "Dead internet theory," et al. Yeah. This has really become a problem. Not for all videos; music videos are kind of fine. I don't listen to music generated by AI but good music should be good music. The rest has unfortunately really gotten worse. Go…

I've seen remarkably little of this when browsing youtube with my cookie (no account, but they know my preferences nonetheless.) Totally different story with a clean fresh session though. One that slipped through, and really pissed me off because it tricked me for a few minutes, was a channel purportedly uploading videos of Richard Feynman explaining things, but the voice and scripts are completely fake. It's disclos…

The barrier to entry for grifting has been lowered, and for existing grifters they can put together some intricate slop. Of course Google doesn't care, they get to show ads against AI slop the same as normal human generated slop.

A fun one was from some minor internet drama around a Battlefield 6 player who seemed to be cheating. A grifter channel pushing some "cheater detection" software started putting out intricate AI generated nonsense that went viral. Searching Karl Jobst CATGIRL will explain.

Re: Google AI Overviews cite YouTube more than any medical site for health queries

#198
post #149
post #138

Earlier quoted context omitted.

> another post-Soviet country Other post-Soviet countries fare substantially better than Russia (Looking at GDP per capita, Russia is about 2500 dollars behind the economic motor of the EU - Bulgaria.)

Must be a misunderstanding 1) Post-soviet countries are doing amazingly well (Poland, Baltics, etc) and very fast growing + healthy (low criminality, etc) 2) The "Russia is weak" thing; it is vastly exaggerated because it is 4 years that we hear that "Russia is on the verge of collapse" but they still manage to handle a very high intensity war against the whole West almost alone. 3) China is not a country lagging beh…

Ukraine is "the whole of the west", interesting? Even the Russian propaganda can't magic up a serious intervention on Ukraine's behalf by western countries. Europeans have been scared to do anything significant, and Trump cut off any real support from the US.

Russia somehow fucked up the initial invasion involving driving a load of preprepared amour across an open border, and have been shredded by FPV drones ever since.

Re: Google AI Overviews cite YouTube more than any medical site for health queries

#199
post #184

Earlier quoted context omitted.

> Not trusting an AI told to lie to you is different than not trusting an AI The entire foundation of trust is that I’m not being lied to. I fail to see a difference. If they are lying, they can’t be trusted

Saying "some people use llms to spread lies therefore I don't trust any llms" is like saying "since people use people to spread lies therefore I don't trust any people". Regardless of whether or not you should trust llms this argument is clearly not proof of it.

Those are false equivalents. If a technology can’t reliably sort out what is a trustworthy source and filter out the rest than it’s not a truth worthy technology. There are tools after all. I should be able to trust a hammer if I use it correctly

All this is also missing the other point: this proves that the narrative companies are selling about AI are not based on objective capabilities

Re: Google AI Overviews cite YouTube more than any medical site for health queries

#200
post #184

Earlier quoted context omitted.

Saying "some people use llms to spread lies therefore I don't trust any llms" is like saying "since people use people to spread lies therefore I don't trust any people". Regardless of whether or not you should trust llms this argument is clearly not proof of it.

Those are false equivalents. If a technology can’t reliably sort out what is a trustworthy source and filter out the rest than it’s not a truth worthy technology. There are tools after all. I should be able to trust a hammer if I use it correctly All this is also missing the other point: this proves that the narrative companies are selling about AI are not based on objective capabilities

The claim here isn't that the technology can't, but that the people using it chose to use it to not. Equivalent to the person with a hammer who chose to smash the 2x4 into pieces instead of driving a nail into it.
Post reply on HN