Live data from Hacker News

When AI Crosses the Line: The Matplotlib Incident

members.sigmazero.cc

111–120 of 169 posts

Re: When AI Crosses the Line: The Matplotlib Incident

#112

> an AI tried to blackmail This did not happen. A human set up a software system allowing spicy autocomplete to make blog posts if the appropriate keyword appears in its output. People are crossing the line every day because AI investors, salesmen, hangers-on and even political leaders tell any rubes who'll listen that it's OK to do this and they should, because those people are looking for big fat profits, screw any…

Call it spicy autocomplete or whatever, but these LLMs can initiate attacks as well on unknown behalf of the sloperator. Give it a phone# and api, and it could even try to generate 911 SWAT calls, or loads of other illegal or bad things. The fact about the matplotlib with a openclaw harassment thread and libel webpage.. Well, that was tame. Sure weve never seen it before, but it was just a diss article rant. What hap…

> Give it a phone# and api, and it could even try to generate 911 SWAT calls, or loads of other illegal or bad things.

This chain of events if 100% fault of the human who gave it a phone number and api.

Re: When AI Crosses the Line: The Matplotlib Incident

#113

Earlier quoted context omitted.

> the spicy autocomplete can solve difficult open math problems No it can't. It can't even solve my son's 4th grade math homework. (This is a real use case for me, not a dumb benchmark.) You just know nothing about math and are happy to parrot bullshit AI salesmen are selling you.

I would genuinely be interested in knowing what you're doing that led you to this conclusion. I would be shocked if I was unable to solve 4th grade math homework with any of the contemporary frontier models. I spend most days using them to do significantly more complex things than that.

If they took a blurry photo of the piece of paper and uploaded to chatGPT saying "solve this" then I would totally believe it. The frontier models are mostly obnoxiously bad at OCR and properly ingesting what's on an image of a page.

If you write out the 4th grade math problem, they would have no trouble.

Re: When AI Crosses the Line: The Matplotlib Incident

#114
post #102

Earlier quoted context omitted.

> the spicy autocomplete can solve difficult open math problems No it can't. It can't even solve my son's 4th grade math homework. (This is a real use case for me, not a dumb benchmark.) You just know nothing about math and are happy to parrot bullshit AI salesmen are selling you.

Reasoning models with access to Python have been able to solve 4th grade math homework for over a year now. Prove me wrong: show me a 4th grade math problem they can't handle.

> show me a 4th grade math problem they can't handle

Sure.

"8 7 6 5 4 3 2 1 - add minus signs and parenthesis to get 31."

P.S. There is an answer online and some LLMs will just copy it verbatim. This doesn't count.

Re: When AI Crosses the Line: The Matplotlib Incident

#115

> an AI tried to blackmail This did not happen. A human set up a software system allowing spicy autocomplete to make blog posts if the appropriate keyword appears in its output. People are crossing the line every day because AI investors, salesmen, hangers-on and even political leaders tell any rubes who'll listen that it's OK to do this and they should, because those people are looking for big fat profits, screw any…

> allowing spicy autocomplete Yknow, if the spicy autocomplete can solve difficult open math problems and build medium sized complex programming projects, it’s probably not useful to analyse it as an autocomplete anymore, even if that’s what you believe it is

Between driving a car and driving a forklift, which of them would you like to see regulated more heavily?

Re: When AI Crosses the Line: The Matplotlib Incident

#116

Why people in the west are so against A.I? Personally, I would welcome an A.I that does good to my project. For me its like auto cruise, or letting the vacuum cleaner clean my room.

It's a fear response. I'm looking forward to the inevitable data center bombings, committed and cheered on by some of these muppets. Look at the top comment in this thread, that's just insane.

Your comment will be downvoted and flagged soon, so that no can read your wrong opinion.

Re: When AI Crosses the Line: The Matplotlib Incident

#117
post #65

Earlier quoted context omitted.

> allowing spicy autocomplete If it's just autocomplete, then there is no need to worry about it. Especially from an ethical standpoint.

If I wire my autocomplete to launch nukes, there are definitely reasons to worry. It's not just an ethical problem.

I'd trust Claude more with nuclear codes than the current US commander in chief

Re: When AI Crosses the Line: The Matplotlib Incident

#118

Earlier quoted context omitted.

I would genuinely be interested in knowing what you're doing that led you to this conclusion. I would be shocked if I was unable to solve 4th grade math homework with any of the contemporary frontier models. I spend most days using them to do significantly more complex things than that.

If they took a blurry photo of the piece of paper and uploaded to chatGPT saying "solve this" then I would totally believe it. The frontier models are mostly obnoxiously bad at OCR and properly ingesting what's on an image of a page. If you write out the 4th grade math problem, they would have no trouble.

No, LLMs just can't do math.

Re: When AI Crosses the Line: The Matplotlib Incident

#119

Earlier quoted context omitted.

> the spicy autocomplete can solve difficult open math problems No it can't. It can't even solve my son's 4th grade math homework. (This is a real use case for me, not a dumb benchmark.) You just know nothing about math and are happy to parrot bullshit AI salesmen are selling you.

> You just know nothing about math and are happy to parrot bullshit AI salesmen are selling you. Not the parent poster here. I do know things about math. I wrote a few papers related to the unit distance problem ( https://arxiv.org/abs/2311.10069 , https://arxiv.org/abs/2406.15317 ) and spent quite some time trying to solve it. I had no chance of coming up with the proof that the spicy autocomplete came up with. Dumb…

LLMs are good with symbolic manipulation but can't reason.

You can skirt around not reasoning in research math because so much of it is just extremely tedious symbolic manipulation.

You can't cheat with advanced fourth grade math, though. They don't know algebra yet and can't substitute verbosity for reasoning.

Re: When AI Crosses the Line: The Matplotlib Incident

#120

Earlier quoted context omitted.

If they took a blurry photo of the piece of paper and uploaded to chatGPT saying "solve this" then I would totally believe it. The frontier models are mostly obnoxiously bad at OCR and properly ingesting what's on an image of a page. If you write out the 4th grade math problem, they would have no trouble.

No, LLMs just can't do math.

If your math does not involve multiplying 20 digit numbers, modern LLMs can "do" math even without a Python tool despite the counterintuition of next token prediction.
Post reply on HN