GPT-5 outperforms federal judges in legal reasoning experiment
141–150 of 254 posts
Re: GPT-5 outperforms federal judges in legal reasoning experiment
#142Earlier quoted context omitted.
Yeah, I'm reminded of the various child porn cases where the "perpetrator" is a stupid teenager who took nude pics of themselves and sent them to their boy/girlfriend. Many of those cases have been struck down by judges because the letter of the law creates a non-sequitur where the teenager is somehow a felon child predator who solely preyed on themselves, and sending them to jail and forcing them to sign up for a se…
Sorry but that seems like an insane system where whole classes of actions effectively are illegal but probably okay if you're likeable. In your scenario the obvious solution is to amend the law and pardon people convinced under it. B/c what really happens is that if you have a pretty face and big tits you get out of speeding tickets b/c "gosh well the law wasn't intended for nice people like you"
https://www.aclu-mn.org/press-releases/victory-judge-dismiss...
"In his decision, Judge Cajacob asserts that the purpose and intent of Minnesota’s child pornography statute does not support punishing Jane Doe for explicit images of herself and doing so “produces an absurd, unreasonable, and unjust result that utterly confounds the statue’s stated purpose.”"
Nothing in there about "likeability" or "we let her off because she had nice tits" (which would be particularly weird in this case). Judges have a degree of discretion to interpret laws, they still have to justify their decisions. If you think the judge is wrong then you can appeal. This is how the law has always worked, and if you've thought otherwise then consider you've been living under this "insane system" for your entire life, and every generation of ancestors has too, assuming you're/they've been in the US.
Re: GPT-5 outperforms federal judges in legal reasoning experiment
#143Earlier quoted context omitted.
I’ve never heard of vigilante justice against someone already sentenced to prison for life, just because they were sentenced in a place without capital punishment? (I mean - people get killed in prison sometimes, I suppose, but it’s not really like vigilante justice on the streets is causing a breakdown in society in Australia, say…)
It's probably rather difficult and risky to enact vigilante justice against someone who's in prison. I think the problem is with places where they don't have life sentences at all, but rather let murderers back out into society after some time. I don't know if vigilante justice is a problem there in reality, but at least I can see it as a possibility: someone might still be angry that you murdered their relative afte…
Having recently done an in-depth review of arguments for and against the death penalty,[1] I can say that this argument is not prominent in the discourse.
Re: GPT-5 outperforms federal judges in legal reasoning experiment
#144Earlier quoted context omitted.
It not the paychecks that influence federal judges; these days it's more of quid-pro-quo for getting the position in the first place. Theoretically they are under no obligation but the bias is built in. The problem with a AI is similar; what in-built biases does it have? Even if it was simply trained on the entire legal history that would bias it towards historical norms.
I think it is usually the opposite - presidents nominate judges they think will agree with them. There’s really nothing a president can do once the judge is sworn in, and we have seen some federal judges take pretty drastic swings in their judicial philosophy over the course of their careers. There’s no reason for the judge to hold up their end of the quid-pro-quo. To the extent they do so, it’s because they were inc…
Re: GPT-5 outperforms federal judges in legal reasoning experiment
#145Earlier quoted context omitted.
Are you even responding to the right comment? I read your comment and the parent comment you've responded to and this response doesn't make sense - it reads like a non-sequitur.
The parent comment present a scenario where the law is ignored b/c the judge decides for himself it shouldn't apply. I'm pointing out that this kind of approach is fundamentally unjust and wrong. "And sure you can say the laws should be written better, but so long as the laws are written by humans that will simply not be the case" The obvious solution is dismissed
Re: GPT-5 outperforms federal judges in legal reasoning experiment
#146Earlier quoted context omitted.
Oddly enough, Texas passed reform to keep sexting teens from getting prosecuted when: they are both under 18 and less than two years difference in age. It was regarded as a model for other states. It's the only positive thing I have heard of Texas legislating wrt sexuality.
> It was regarded as a model for other states. Really? That "model" has the common, but obviously extremely undesirable, feature of criminalizing sexual relationships between students in the same grade that were legal when they formed . How could it be regarded as a model for anyone else?
Re: GPT-5 outperforms federal judges in legal reasoning experiment
#147IANAL, but this seems like an odd test to me. Judges do what their name implies - make judgment calls. I find it re-assuring that judges get different answers under different scenarios, because it means they are listening and making judgment calls. If LLMs give only one answer, no matter what nuances are at play, that sounds like they are failing to judge and instead are diminishing the thought process down to black-…
There is no rule that can be written so precisely that there are no exceptions, including this one.
A joke[0], but one I think people should take seriously. Law would be easy if it weren't for all the edge cases. Most of the things in the world would be easy if it weren't for all the edge cases[1]. This can be seen just by contemplating whatever domain you feel you have achieved mastery over and have worked with for years. You likely don't actually feel you have achieved mastery because you're developed to the point where you know there is so much you don't know[2].The reason I wouldn't want an LLM judge (or any algorithmic judge) is the same reason I despise bureaucracy. Bureaucracy fucks everything up because it makes the naive assumption that you can figure everything out from a spreadsheet. It is the equivalent of trying to plan a city from the view out of an airplane window. The perspective has some utility, but it is also disconnected from reality.
I'd also say that this feature of the world is part of what created us and made us the way we are. Humans are so successful because of our adaptability. If this wasn't a useful feature we'd have become far more robotic because it would be a much easier thing for biology to optimize. So when people say bureaucracies are dehumanizing, I take it quite literally. There's utility to it, but its utility leads to its overuse and the bias is clear that it is much harder to "de"-implement something than to implement it. We should strongly consider that bias in society when making large decisions like implementing algorithmic judges. I'm sure they can be helpful in the courtroom, but to abdicate our judgements to them only results in a dehumanized justice system. There are multiple literal interpretations of that claim too.
[0] You didn't look at my name, did you?
[1] https://news.ycombinator.com/item?id=43087779
[2] Hell, I have a PhD and I forget I'm an expert in my domain because there's just so much I don't know I continue to feel pretty dumb (which is also a driving force to continue learning).
Re: GPT-5 outperforms federal judges in legal reasoning experiment
#148IANAL, but this seems like an odd test to me. Judges do what their name implies - make judgment calls. I find it re-assuring that judges get different answers under different scenarios, because it means they are listening and making judgment calls. If LLMs give only one answer, no matter what nuances are at play, that sounds like they are failing to judge and instead are diminishing the thought process down to black-…
The main job of a judicial system is to appear just to people. As long as people think it's just -- everyone is happy. But if it's strictly by the law, but people consider it's unjust -- revolutions happen. In both cases, lawmakers must adapt the law to reflect what people think is "just". That's why there are jury duty in some countries -- to involve people to the ruling, so they see it's just .
> to appear just to people.
The best way to appear just is to be just.But I'm not sure what your argument is. It is our duty as citizens to encourage the system to be just. Since there is no concrete mathematical objective definition of justice, well, then... all we can work with is the appearance. So I don't think your insight is so much based on some diabolical deep state thinking but more on the limitations of practicality. Your thesis holds true if everyone is trying their best to be just.
Re: GPT-5 outperforms federal judges in legal reasoning experiment
#149Count me out of a society that uses LLMs to make rulings. The dystopia of having to find a lawyer who is best at promoting the "unbiased" judge sounds like a hellscape.