Live data from Hacker News

Lawyer cites fake cases invented by ChatGPT, judge is not amused

simonwillison.net

81–90 of 319 posts

Re: Lawyer cites fake cases invented by ChatGPT, judge is not amused

#81
post #7

No, it did not “double-check”—that’s not something it can do! And stating that the cases “can be found on legal research databases” is a flat out lie. What’s harder is explaining why ChatGPT would lie in this way. What possible reason could LLM companies have for shipping a model that does this? It did this because it's copying how humans talk, not what humans do. Humans say "I double checked" when asked to verify so…

ChatGPT did not lie; it cannot lie. It was given a sequence of words and tasked with producing a subsequent sequence of words that satisfy with high probability the constraints of the model. It did that admirably. It's not its fault, or in my opinion OpenAI's fault, that the output is being misunderstood and misused by people who can't be bothered understanding it and project their own ideas of how it should function…

>ChatGPT did not lie; it cannot lie.

If it lies like a duck, it is a lying duck.

Re: Lawyer cites fake cases invented by ChatGPT, judge is not amused

#83

I went ahead and asked ChatGPT with the browsing plugin [1] because I was curious and it answered that it was a real case citing an article about the fake citations! After some prodding ("Are you sure?") it spat out something slightly saner citing this very article! > The case "Varghese v. China Southern Airlines Co., Ltd., 925 F.3d 1339 (11th Cir. 2019)" was cited in court documents, but it appears that there might…

In the loop there indeed was a allegedly trained human in this instance

That's not what I would call in the loop. He didn't check that the sources were real.

By "in the loop" I mean actively validating statements of fact generated by ChatGPT

Re: Lawyer cites fake cases invented by ChatGPT, judge is not amused

#84
post #40

Earlier quoted context omitted.

"It doesn't lie, it just generates lies and printed them to the screen!" I don't think there's a difference.

Saying ChatGPT lies is like saying The Onion lies.

The Onion (via its staff) intends to produce falsehoods. ChatGPT (nor its staff) does not.

Re: Lawyer cites fake cases invented by ChatGPT, judge is not amused

#85

Earlier quoted context omitted.

I wonder if this is a tactic so the court to deems this lawyer incompetent rather than giving the (presumably much harsher) penalty for deliberately lying to the court?

Why assume malice? Asking ChatGPT to verify is exactly what someone who trusts ChatGPT might do. I'm not surprised this lawyer trusted ChatGPT too much. People trust their lives to self driving cars, trust their businesses to AI risk models, trust criminal prosecution to facial recognition. People outside the AI field seem to be either far too trusting or far too suspicious of AI.

Quoted directly from my last session with ChatGPT mere seconds ago:

> Limitations

May occasionally generate incorrect information

May occasionally produce harmful instructions or biased content

Limited knowledge of world and events after 2021

---

A lawyer who isn't prepared to read and heed the very obvious warnings at the start of every ChatGPT chat isn't worth a briefcase of empty promises.

WARNING: witty ending of previous sentence written with help from ChatGPT.

Re: Lawyer cites fake cases invented by ChatGPT, judge is not amused

#87
post #40

Earlier quoted context omitted.

"It doesn't lie, it just generates lies and printed them to the screen!" I don't think there's a difference.

The difference is everything. It doesn't understand intent, it doesn't have a motivation. This is no different than what fiction authors, songwriters, poets and painters do. The fact that people assume what it produces must always be real because it is sometimes real is not its fault. That lies with the people who uncritically accept what they are told.

> That lies with the people who uncritically accept what they are told.

That's partly true. Just as much fault lies with the people who market it as "intelligence" to those who uncritically accept what they are told.

Re: Lawyer cites fake cases invented by ChatGPT, judge is not amused

#88
post #40

Earlier quoted context omitted.

"It doesn't lie, it just generates lies and printed them to the screen!" I don't think there's a difference.

To perhaps stir the "what do words really mean" argument, "lying" would generally imply some sort of conscious intent to bend or break the truth. A language model is not consciously making decisions about what to say, it is statistically choosing words which probabilistically sound "good" together.

>A language model is not consciously making decisions about what to say

Well, that is being doubted -- and by some of the biggest names in the field.

Namely that it isn't "statistically choosing words which probabilistically sound good together". But that doing so is not already making a consciousness (even if basic) emerge.

>it is statistically choosing words which probabilistically sound "good" together.

That when we do speak (or lie), we do something much more nuanced, and not just do a higher level equivalent of the same thing, plus have the emergent illusion of consciousness, is also an idea thrown around.

Re: Lawyer cites fake cases invented by ChatGPT, judge is not amused

#89

This is why it is very important to have the prompts fill in relevant fragments from a quality corpus. That people think these models “tell the truth” or “hallucinate” is only half the story. It’s like expecting your language center to know all the facts your visual consciousness contains, or your visual consciousness to be able to talk in full sentences. It’s only when all models are working well together the truth…

[deleted]

Re: Lawyer cites fake cases invented by ChatGPT, judge is not amused

#90

Earlier quoted context omitted.

GPT4 can double-check to an extent. I gave it a sequence of 67 letter As and asked it to count them. It said "100", I said "recount": 98, recount, 69, recount, 67, recount, 67, recount, 67, recount, 67. It converged to the correct count and stayed there. This is quite a different scenario though, tangential to your [correct] point.

The example of asking it things like counting or sequences isn't a great one because it's been solved by asking it to "translate" to code and then run the code. I took this up as a challenge a while back with a similar line of reasoning on Reddit (that it couldn't do such a thing) and ended up implementing it in my AI web shell thing. heavy-magpie|> I am feeling excited. system=> History has been loaded. pastel-matur…

Shouldn’t the answer be zero?
Post reply on HN