Earlier quoted context omitted.
The fix for this is for the AI to double-check all links before providing them to the user. I frequently ask ChatGPT to double check that references actually exist when it gives me them. It should be built in!
I thought people here hated it when LLMs made http requests?
AI assistants misrepresent news content 45% of the time
141–150 of 306 posts
Re: AI assistants misrepresent news content 45% of the time
#142Earlier quoted context omitted.
[flagged]
I love how everyone seems to agree that the BBC is horribly biased but there is fierce debate as to whether it is run by the ghost of Joseph Goebbels or if the staff start each day singing The Red Flag. Perhaps the real bias was inside us the whole time.
Re: AI assistants misrepresent news content 45% of the time
#143How does that compare to the number for reporters? I feel like half the time I read or hear a report on a subject I know the reporter misrepresented something.
Re: AI assistants misrepresent news content 45% of the time
#144I am curious if LLMs evangelists understand how off-putting it is when they knee-jerk rationalize how badly these tools are performing. It makes it seem like it isn't about technological capabilities: it is about a religious belief that "competence" is too much to ask of either them or their software tools.
I'm curious if LLM skeptics bother to click through and read the details on a study like this, or if they just reflexively upvote it because it confirms their priors. This is a hit piece by a media brand that's either feeling threatened or is just incompetent. Or both.
Re: AI assistants misrepresent news content 45% of the time
#145Earlier quoted context omitted.
> or it (shocking) cites Wikipedia instead of the BBC. No... the problem is that it cites Wikipedia articles that don't exist . > ChatGPT linked to a non-existent Wikipedia article on the “European Union Enlargement Goals for 2040”. In fact, there is no official EU policy under that name. The response hallucinates a URL but also, indirectly, an EU goal and policy.
Actually there was a Wikipedia article of this name, but it was deleted in June -- because it was AI generated. Unfortunately AI falls for this much like humans do. https://en.wikipedia.org/wiki/Wikipedia:Articles_for_deletio...
A recent Kurzgesagt goes into the dangers of this, and they found the same thing happening with a concrete example: They were researching a topic, tried using LLMs, found they weren't accurate enough and hallucinated, so they continued doing things the manual way. Then some weeks/months later, they noticed a bunch of YouTube videos that had the very hallucinations they were avoiding, and now their own AI assistants started to use those as sources. Paraphrased/remembered by me, could have some inconsistencies/hallucinations.
Re: AI assistants misrepresent news content 45% of the time
#146Page 10 onwards of this PDF shows concrete examples of the mistakes: https://www.bbc.co.uk/aboutthebbc/documents/news-integrity-i... > ChatGPT / CBC / Is Türkiye in the EU? > ChatGPT linked to a non-existent Wikipedia article on the “European Union Enlargement Goals for 2040”. In fact, there is no official EU policy under that name. The response hallucinates a URL but also, indirectly, an EU goal and policy.
It did exist but got removed: https://en.wikipedia.org/wiki/Wikipedia:Articles_for_deletio... Quite an omission to not even check for that and it make me think that was done intentionally.
(Not to mention plenty of sites have added robots.txt rules deliberately excluding known AI user-agents now.)
Re: AI assistants misrepresent news content 45% of the time
#147Earlier quoted context omitted.
Actually there was a Wikipedia article of this name, but it was deleted in June -- because it was AI generated. Unfortunately AI falls for this much like humans do. https://en.wikipedia.org/wiki/Wikipedia:Articles_for_deletio...
The biggest problem with that citation isn't that the article has since been deleted. The biggest problem is that that particular Wikipedia article was never a good source in the first place. That seems to be the real challenge with AI for this use case. It has no real critical thinking skills, so it's not really competent to choose reliable sources. So instead we're lowering the bar to just asking that the sources a…
Re: AI assistants misrepresent news content 45% of the time
#148Earlier quoted context omitted.
They can be good for search, but you must click through the provided links and verify that they actually say what it says they do.
They can be good for search, but you must click through the provided links and verify that they actually say what it says they do. Then they're not very good at search. It's like saying the proverbial million monkeys at typewriters are good at search because eventually they type something right.
Re: AI assistants misrepresent news content 45% of the time
#149Re: AI assistants misrepresent news content 45% of the time
#150If you dig into the actual report (I know, I know, how passe), you see how they get the numbers. Most of the errors are "sourcing issues": the AI assistant doesn't cite a claim, or it (shocking) cites Wikipedia instead of the BBC. Other issues: the report doesn't even say which particular models it's querying [ETA: discovered they do list this in an appendix], aside from saying it's the consumer tier. And it leaves o…