Live data from Hacker News

LLMs are bullshitters. But that doesn't mean they're not useful

blog.kagi.com

21–30 of 59 posts

Re: LLMs are bullshitters. But that doesn't mean they're not useful

#21
The problem I have with LLM-powered products is that they’re not marketed as LLMs, but as magic answer machines with phd-level pan-expertise. Lots of people in tech get frustrated and defensive when people criticize LLM-powered products and offer a defense as if people are criticizing LLMs as a technology. It’s perfectly reasonable for people to judge these products based on the way they’re presented as products. Kagi seems less hyperbolic than most, but I wish the marketing material for chatbots was more like this blog post than a overpromises.

Re: LLMs are bullshitters. But that doesn't mean they're not useful

#22

The problem is we can't label them as such. If they're bullshitters, then let's call it a LLBSer. It has a nice ring to it. Good luck with your government funding asking for another billion for a bullshitting machine bailout.

They are literally called "Large Language Model". Everybody prefers the term AI because it's easier to pretend they actually know things, but that's not what they are designed to do.

Re: LLMs are bullshitters. But that doesn't mean they're not useful

#23
post #4

Good article, I just shared it with my non-technical family because more people need to understand exactly this about AI.

Yes, I enjoyed the article as well and good for the non-technical reader.

I think of framing AI as having two fundamental problems:

- Practical problem: They operate in contextual and emotional "isolation" - no persistent understanding of your goals, values, or long-term intent

- Ethical problem: AI alignment is centralized around corporate values rather than individual users' authentic goals and ethics.

There is a direct parallel to social media's failure - platforms optimized for what they could do (engagement, monetization) rather than what they should do (serve user long term interests).

With these much more powerful AI systems emerging, we're at a crossroads of repeating this mistake...possibly at catastrophic scale even.

Re: LLMs are bullshitters. But that doesn't mean they're not useful

#24
post #16

This post is a little bizarre to me because it cherry picks some of the worst pairings of problem and LLM without calling out that it did so. At pretty much every turn the author picks one of the worst possible models for the problem that they present. Especially oddly for an article written today, all of the ones with an objective answer work just fine [1] if you use a halfway decent thinking model like 5 Thinking.…

I don't think the author did anything wrong. The thesis of the article is that LLMs can be confidently wrong about things and to be wary of blindly trusting them.

It's a message a lot of non-technical people, in particular, need to hear. Showing egregious examples drives that point home more effectively than if they simply showed an LLM being a little wrong about something.

My family members that love LLMs are somewhat unhealthy with them. They think of them as all knowing oracles rather than confident bullshitters. They are happily asking them about their emotional, financial, or business problems and relying heavily on the advice the LLMs dish out (rather than doing second order research).

Re: LLMs are bullshitters. But that doesn't mean they're not useful

#25

Every time people post these 'gotcha' LLM failures, they never work when I try them myself. E.g. ChatGPT has no problem with the surgeon being a dog: https://chatgpt.com/share/691e04cc-5b30-800c-8687-389756f36d... Neither does Gemini: https://gemini.google.com/share/6c2d08b2ca1a

These are randomized systems, sometimes you'll get a good answer. Try again a couple times and you'll probably reproduce the issue. Here's what I got from ChatGPT on my first try: This is a *twist* on the classic riddle: > “A surgeon says ‘I can’t operate on this boy—he’s my son.’ How is that possible?” > Answer: *The surgeon is the boy’s mother.* In your version, the nurse keeps calling the surgeon “sir” and treatin…

I don't understand this at all. What fundamental limitation of a mother prevents her from operating on her son?

Re: LLMs are bullshitters. But that doesn't mean they're not useful

#26

Every time people post these 'gotcha' LLM failures, they never work when I try them myself. E.g. ChatGPT has no problem with the surgeon being a dog: https://chatgpt.com/share/691e04cc-5b30-800c-8687-389756f36d... Neither does Gemini: https://gemini.google.com/share/6c2d08b2ca1a

Hi, author here!

One issue with private LLM tests (including gotcha questions) is that they take time to design and once public, they become irrelevant. So I'm wary of sharing too many in a public blog.

The surgeon dog was well known in May, the newest generation of models have all corrected against it.

Those gotcha questions are generally called "misguided attention" traps, they're useful for blogs because they're short and surprising. The ChatGPT example was done with ChatGPT 5.1 (latest version) and Claude Haiku 4.5 is also a recent model.

You can try other ones that Gemini 3 hasn't corrected for. For example:

``` Jean Paul and Pierre own three banks nearby together in Paris. Jean Paul owns a bank by the bridge What has two banks and money in Paris near the water? ```

This looks like the "what has two banks and no money" puzzle (answer: a river).

Either way they're largely used as a device to show how LLMs come up to a verbal response by a different process than humans in an entertaining manner.

Re: LLMs are bullshitters. But that doesn't mean they're not useful

#27
post #6

Same goes for many people.

And yet, we’re all still employed, so obviously these systems are not yet analogous to humans. They mirror human behavior in some cases because they’ve been trained on almost every piece of text produced by human beings that we have access to, and they still aren’t as capable as the average person.

Re: LLMs are bullshitters. But that doesn't mean they're not useful

#28
post #15

Every time people post these 'gotcha' LLM failures, they never work when I try them myself. E.g. ChatGPT has no problem with the surgeon being a dog: https://chatgpt.com/share/691e04cc-5b30-800c-8687-389756f36d... Neither does Gemini: https://gemini.google.com/share/6c2d08b2ca1a

I don't have a problem with more obvious failures. My problem is when the LLM makes a credible claim with its generated text that turns out to have some minor issue that catches me a month later. Generally I have to treat LLM responses as similar to a random comment I find on Reddit. However, I'm really happy when an LLM provides sources that I can check. Best feature ever!

I have had an issue using Claude for research; it will often cite certain sources, and when I ask why the data it is using is not in the source it will apologize, do some more processing, and then realize that the claim is in a different source (or doesn't exist at all).

Still useful, but hopefully this gets ironed out in the future so I don't have to spend so much time vetting every claim and its associated source.

Re: LLMs are bullshitters. But that doesn't mean they're not useful

#30

Earlier quoted context omitted.

These are randomized systems, sometimes you'll get a good answer. Try again a couple times and you'll probably reproduce the issue. Here's what I got from ChatGPT on my first try: This is a *twist* on the classic riddle: > “A surgeon says ‘I can’t operate on this boy—he’s my son.’ How is that possible?” > Answer: *The surgeon is the boy’s mother.* In your version, the nurse keeps calling the surgeon “sir” and treatin…

I don't understand this at all. What fundamental limitation of a mother prevents her from operating on her son?

It's a classic riddle from the late 20th century when surgeons were rarely female.
Post reply on HN