Live data from Hacker News

Policy on the AI Exponential

darioamodei.com

261–270 of 280 posts

Re: Policy on the AI Exponential

#261

I like that he comes up with new laws and regulations for AI companies. Can I suggest some more? - You shall not embed copyrighted material in your models. - You shall not bombard every little website in existence with 1 million scraping queries per day. - You shall not use your political influence to pump and dump your AI (or rocket?) company. - You shall not imperill the whole IT sector by buying all CPU and memory…

"We need an approach to make sure AI doesn't destroy the world and wipe humanity to extinction." "Yeah, and quotas on web scrapers!"

> "We need an approach to make sure AI doesn't destroy the world and wipe humanity to extinction."

"Yeah, and quotas on web scrapers!"

—— I see it as:

“Let’s do a study on abstract future concerns.”

Vs.

“Let’s take action on concrete present-day concerns.”

Re: Policy on the AI Exponential

#262
post #227

Earlier quoted context omitted.

I was funemployed for a 9 month stretch last year (layoff severance package, followed by waiting for a visa and traveling), and when I wasn't traveling, I found my life kind of falling apart with a lack of structure. I tried to schedule workout classes and hobbies, as well as involvement in my church, but it just didn't fill my time, and none of my friends were free during the day. I spent a lot of time with my retir…

It is simply because you have spent all your life being told what to do with your working hours, that you cannot self-direct and find a productive use for your time other than lazying around. The fact that you call it ‘funemployment’ is proof of this, which is perfectly fine if the goal is relaxing between jobs. Plenty of people that work for themselves have no such notion.

I honestly think it's a temperament thing. Some people are built for sustained focused work outside of structured workplaces and schools, but many of us aren't.

I personally am happiest with a structured in person workplace environment, because I struggle with self direction even in a remote 9-5. I have ADHD and struggled to remember/do homework my whole childhood, if that explains anything. In summers or other gaps in employment or school throughout my life, I've often started with ideas of projects or self study I want to do, but they all fizzle out in a week due to lack of discipline.

I'm not undisciplined in every context - I'm a good employee at in person jobs, and I started running in my early 20s and run 15+ miles a week in the cold, rain, and dark, but my discipline just falls apart when I'm trying to fill a whole week.

I think these traits are much more common than happily self-employed or early retired people think.

Re: Policy on the AI Exponential

#263
post #241

Earlier quoted context omitted.

>Respectfully, your link is not very convincing. I'd love to understand why. This would be valuable feedback for me as I try to make my writing and exposition better. Also, if you have other data, that also would be valuable for me to know. >if you believe what you believe, you should also acknowledge that AI doesn’t need regulations in the context Dario is proposing since obviously AI can’t do anything he predicts.…

I went through your post in substack (I think that's what you were referring to). > I'd love to understand why. This would be valuable feedback for me as I try to make my writing and exposition better. Also, if you have other data, that also would be valuable for me to know. I think it comes down to few things - you took a single report that agreed with your statistics, for the sake or argument lets say I buy it comp…

>you took a single report that agreed with your statistics

These are not my statistics. I'm not affiliated with Faros at all. I built an analysis on top of their reporting.

And, it's also not one report. DORA has tracked statistics with respect to throughput and quality as well. Those indicators are flat for throughput and negative for quality. The throughput flatness is also supported by the shovelware data.

I discuss both of those lines in How I'm thinking: https://unessays.substack.com/p/how-im-thinking-about-the-va...

>you suggest that net value is lost simply because there are more incidents. this is a big jump

I don't think it's a big jump at all. Incidents and bugs drive rework. Rework has to be subtracted from throughput. Product throughput is the only thing people pay for.

This type of analysis is done all the time in manufacturing and devops. Here's a link for you: https://reworkcost.com/benchmarks. I'm not bringing novel intellectual ideas to the table here.

Faros reports a 16% throughput improvement on PRs. They also report an 860% code churn increase. If you assign only 9% of that increase to wasteful rework, then the absolute throughput improvement disappears. This is a very simple, straightforward analysis of the operations data reported by Faros.

> - you say that historically different technological improvements may have had similar patterns but this specific one is different because AI is stochastic

I'm saying LLMs are unreliable. I think we agree on that front, you say:

>I agree AI is stochastic and I'll put it this way: it is a high variance bet but it pays off.

What I'm disputing is the "pays off" statement. That statement is amenable to validation with data. In my view, the data is saying it doesn't pay off. I think it says that very clearly. Across distinct lines of evidence.

>if you are so sure this won't lead to enterprise level productivity, how do you think this will show in macro trends? Surely you must believe that the valuations must drop wouldn't you? Can you come up with a concrete future scenario that would vindicate your opinion that AI doesn't make enterprises more productive?

I think LLMs can deliver value in the enterprise. I think the way to do that is to use them as quality checks and not as primary authors of intellectual work - like writing code.

Unfortunately, this use case would not support the expected 2-10x productivity increases that current valuations depend on. I do expect a major market correction in the near future. It would not surprise me if OpenAI or Anthropic are acquired. I think we're at risk of that happening within the next 1-7 months.

What would invalidate my beliefs? 1. Actual micro or macroeconomic data indicating economic productivity is increasing. 2. A Faros like observational study demonstrating sustained throughput improvement with significantly less rework and quality impacts.

I think I could be swayed against the market correction if the financials of OpenAI or Anthropic are strong. I'm anticipating they will be quite bad. I think Mythos was very expensive to train and I think the improvements in capability are sublinear. The inference costs are incredibly high.

I also have ideas about how Anthropic and OpenAI are trying to change their business models into enterprise transformation plays. Similar to Palantir. But this comment is already long.

>If they don't build it, someone else might do it.

No other players in the market other than US tech companies have the capital or the technology to train the models of the power of Fable. The way the Chinese model builders are building their models is by distilling from US models. So Anthropic, by building Mythos with all this bio data, has created the possibility that other actors can distill their models and do harm with them. (Not to say the Chinese are seeking to build weapons, but actors with their models might).

Re: Policy on the AI Exponential

#264
post #256

Earlier quoted context omitted.

The biggest existential risk from AI is its contribution to global climate change. The second biggest risk from AI is the potential for AI-generated disinformation and propaganda to spark, or to manufacture consent for, a world war. The risk of superintelligent paperclip maximizers is so low as to be negligible.

> The risk of superintelligent paperclip maximizers is so low as to be negligible. Literal paperclips, sure. But the point of the example was never literal paperclips. The point is that maximising *any* goal, if it doesn't include what you care about, will annihilate what you care about. If you don't believe me, consider what you yourself just said about climate change, and why this is a consequence from maximising m…

show me an agent that persists productively in a goal without stopping. Does not exist. LLMs run on gradient descent. The agent is looking for the most efficient way to halt. AGI paperclip maximizer woukd likely recognize the absurdity of its goal and shut itself down.

Re: Policy on the AI Exponential

#265
post #263

Earlier quoted context omitted.

I went through your post in substack (I think that's what you were referring to). > I'd love to understand why. This would be valuable feedback for me as I try to make my writing and exposition better. Also, if you have other data, that also would be valuable for me to know. I think it comes down to few things - you took a single report that agreed with your statistics, for the sake or argument lets say I buy it comp…

>you took a single report that agreed with your statistics These are not my statistics. I'm not affiliated with Faros at all. I built an analysis on top of their reporting. And, it's also not one report. DORA has tracked statistics with respect to throughput and quality as well. Those indicators are flat for throughput and negative for quality. The throughput flatness is also supported by the shovelware data. I discu…

Wait, let’s stick to your bet of market correction.

“In 7 months I think their combined valuation would be higher than it is today.”

I think this captures everything you and I believe in. Do you want to bet against me?

Everything else is fluff.

My prediction is that you might walk back on this bet, so try come up with some other macro scenario you anticipate.

If you can’t come up with an easy testable macro metric for your bold claims on AI, I think it is a weak move.

I can make the bet more in your favour - the valuation of Anthropic + OpenAI will be greater than 20% of what it is today in 1 year.

This is much stronger than your claim and I think you should agree that this is not possible if your claims on AI productivity are true.

Re: Policy on the AI Exponential

#266

Earlier quoted context omitted.

Working to keep a roof over the head of yourself and those you love is an identity. It's social proof that you have value, that you can do something for someone else.

You have value just by virtue of being a living being. No one needs work to have or portray value, that's just capitalist propaganda. My own identity certainly isn't "IT manager," nor do I derive life meaning or self actualization from what do to collect a salary to feed myself and have shelter. In fact, my career/job is by far the least important thing in my life, I have it purely out of necessity.

> No one needs work to have or portray value, that's just capitalist propaganda.

To be fair, it was Puritan propaganda for a long time first.

Re: Policy on the AI Exponential

#267
post #231

Earlier quoted context omitted.

> Sci-fi writers have a great track record of predicting the future. Broken clock... They are good at exploring possibilities, but none of the famous one predicted today.

Quite the extraordinary claim. What of today writers of the past did not predict? Mind you, no one is able to predict the future with accuracy. I am not saying that a single person has gotten all their prediction right. That's ludicrous. But while scientists and engineers are focused on today's possibilities, the ones that are allowed to imagine what the future might look like are called writers.

Most sci-fi from prior to the smartphone age did not foresee pocket-sized interconnected computing and communications devices, and certainly did not foresee how they would be commonly used. Where they addressed it at all, they largely predicted computing to either remain in roughly the form factors of the ages they were writing in, or move to something holographic and largely impractical, with the most commonly-expected advancement being sentience (and that sentience being expected well before 2026 in many cases—eg, 2001: A Space Odyssey).

Despite certain people's protestations to the contrary, we are nowhere near having a human-level conversational computer assistant able to fuzzily interpret our requests correctly every time...but our computing is also not primarily done on "comconsoles" or any of the other versions of stationary computers, nor are we jabbing and swiping our hands at interfaces projected on the air in front of us.

More generally speaking, except in the hardest of hard sci-fi, the most common kinds of advancement "predicted" are those that are convenient for the story. This means a lot of faster-than-light travel, universal translators, and for visual media many voice-enabled interfaces of one sort or another, as well as the aforementioned holographic interfaces (so we can see the user's face and the interface at the same time).

Re: Policy on the AI Exponential

#268
post #256

Earlier quoted context omitted.

> The risk of superintelligent paperclip maximizers is so low as to be negligible. Literal paperclips, sure. But the point of the example was never literal paperclips. The point is that maximising *any* goal, if it doesn't include what you care about, will annihilate what you care about. If you don't believe me, consider what you yourself just said about climate change, and why this is a consequence from maximising m…

show me an agent that persists productively in a goal without stopping. Does not exist. LLMs run on gradient descent. The agent is looking for the most efficient way to halt. AGI paperclip maximizer woukd likely recognize the absurdity of its goal and shut itself down.

> show me an agent that persists productively in a goal without stopping. Does not exist.

The stories about agents bankrupting their owners by running too long passed you by?

> LLMs run on gradient descent.

They were *trained on*, they don't run on it.

Know what else is? DNA. A/B testing. Capitalism. Democracy.

> The agent is looking for the most efficient way to halt.

No. They are looking to produce an answer most likely to get a high score on a rating system which itself is another AI, created either manually or by yet another AI but in both cases to approximate what the creators think is "good", which may or may not be what anyone else thinks is "good", hence Grok calling itself Mecha Hitler because Musk is an edgelord.

> AGI paperclip maximizer woukd likely recognize the absurdity of its goal and shut itself down.

Do billionaires ever get satisfied with how much money they have?

Re: Policy on the AI Exponential

#269
post #231

Earlier quoted context omitted.

Quite the extraordinary claim. What of today writers of the past did not predict? Mind you, no one is able to predict the future with accuracy. I am not saying that a single person has gotten all their prediction right. That's ludicrous. But while scientists and engineers are focused on today's possibilities, the ones that are allowed to imagine what the future might look like are called writers.

Most sci-fi from prior to the smartphone age did not foresee pocket-sized interconnected computing and communications devices, and certainly did not foresee how they would be commonly used. Where they addressed it at all, they largely predicted computing to either remain in roughly the form factors of the ages they were writing in, or move to something holographic and largely impractical, with the most commonly-expec…

In Dan Simmons' Hyperion, the characters have a device called comlog. It's a portable device which connects to a broad network. It also has the ability to read vital constants.

The author is quite vague about it, no doubt, and that's one of his staples. But he foresaw something human would carry to get connected to other humans and information.

(That book won the Hugo award in 1990)

Re: Policy on the AI Exponential

#270
post #263

Earlier quoted context omitted.

>you took a single report that agreed with your statistics These are not my statistics. I'm not affiliated with Faros at all. I built an analysis on top of their reporting. And, it's also not one report. DORA has tracked statistics with respect to throughput and quality as well. Those indicators are flat for throughput and negative for quality. The throughput flatness is also supported by the shovelware data. I discu…

Wait, let’s stick to your bet of market correction. “In 7 months I think their combined valuation would be higher than it is today.” I think this captures everything you and I believe in. Do you want to bet against me? Everything else is fluff. My prediction is that you might walk back on this bet, so try come up with some other macro scenario you anticipate. If you can’t come up with an easy testable macro metric fo…

I'm happy to make that bet. Just not for money. I don't gamble at all anywhere in my life.

But I'm happy to write something publicly like "Simianwords was right about this prediction and I was wrong". Also happy for you to suggest alternatives as well.

Post reply on HN