A Man Out to Prove How Dumb AI Still Is
theatlantic.com
A Man Out to Prove How Dumb AI Still Is
1–10 of 66 posts
Re: A Man Out to Prove How Dumb AI Still Is
#2s/intellectually lazy/hype maxing for fundraising/
Re: A Man Out to Prove How Dumb AI Still Is
#3Re: A Man Out to Prove How Dumb AI Still Is
#4> the bot came up with more than 1,000 possible answers per grid before selecting a final submission.
Yeah, AGI is right around the corner… /s
Re: A Man Out to Prove How Dumb AI Still Is
#5I feel like this description really buries the lede on Chollet's expertise. (For those who don't know, he's the creator of and lead contributor[0] to Keras)
Re: A Man Out to Prove How Dumb AI Still Is
#6> When I spoke with him earlier this year, Chollet told me that AI companies have long been “intellectually lazy“ s/intellectually lazy/hype maxing for fundraising/
Re: A Man Out to Prove How Dumb AI Still Is
#7Re: A Man Out to Prove How Dumb AI Still Is
#8> When I spoke with him earlier this year, Chollet told me that AI companies have long been “intellectually lazy“ s/intellectually lazy/hype maxing for fundraising/
I think it's fascinating that his impossible benchmark got defeated, but because the Keras guy doesn't like LLMs, it is possible to mishear algorithmic distaste as saying people shipping this are "lazy" and "hype maxing."
Re: A Man Out to Prove How Dumb AI Still Is
#9Arc AGI is the main reason why I don't trust static bench marks.
If you don't have an essentially infinite set to draw your validation data from then a large enough model will memorize it as part of its developer teams KPIs.
Forget all these fancy benchmarks. If you want to saturate any model today give it a string and a grammar and ask it to generate the string from the grammar. I've had _every_ model fail this on regular grammars with strings of more than 4 characters long.
LLMs are the solution to natural language, which is a huge deal. They aren't the solution to reasoning which is still best solved with what used to be called symbolic AI before it started working, e.g. sat solvers.
Re: A Man Out to Prove How Dumb AI Still Is
#10Seems to me that a lot of folks are enjoying having an LLM rewrite their email or whatever, but I wonder how many are actually buying the rest of it? The companies themselves sure aren’t helping.