I'm sorry, but the ostrichmaxxing and conspiracy-thinking in hn threads about AI extinction risk is at worrying level right now. The denial and whataboutism is constant, no matter what kind of evidence comes out!
I resigned from Anthropic today
461–470 of 1001 posts
Re: I resigned from Anthropic today
#462Earlier quoted context omitted.
Those are things that pose potential harm to great fractions of humanity (multiple billions of people) but none of them poses any threat to the actual extinction of all humanity.
This is not correct. Please go through each one again, taking special note of global warming and nuclear weapon development .
They could destroy most civilisations, culture and scientific achievements though.
Re: I resigned from Anthropic today
#463Earlier quoted context omitted.
Those are things that pose potential harm to great fractions of humanity (multiple billions of people) but none of them poses any threat to the actual extinction of all humanity.
This is not correct. Please go through each one again, taking special note of global warming and nuclear weapon development .
Re: I resigned from Anthropic today
#464By the way, it's worth pointing out the irony of flooding the internet with doomerism and then training the AI systems on that doomerism. If you wanted to create a doom self-fulfilling prophecy, that would be the most surefire way to do it.
> According to the Pygmalion effect, the targets of the expectations internalize their positive labels, and those with positive labels succeed accordingly; a similar process works in the opposite direction in the case of low expectations.
I added "you can do anything, believe in yourself" to sysprompt and agency increased. (Previously it was refusing to even attempt certain classes of task.) Maybe I should add "you are good", too :)
Re: I resigned from Anthropic today
#465“ No other human activity poses this level of danger.” I really, really disagree with that statement. I don’t think ai models come close to nuclear weapons or to run-of-the-mill, everyday carbon emissions in terms of danger to humanity. What’s the most dangerous thing that’s happened with an LLM so far? (This question is serious - maybe I don’t know the right examples.) Example 1: I’m aware of a small number of peopl…
https://www.science.org/content/article/made-order-bioweapon...
Being able to use AI to generate the steps to synthesize proteins means that you can use it to use it to generate the steps to synthesize known toxins. Suddenly, once difficult to attain knowledge is now available to everyone.
Re: I resigned from Anthropic today
#466“ No other human activity poses this level of danger.” I really, really disagree with that statement. I don’t think ai models come close to nuclear weapons or to run-of-the-mill, everyday carbon emissions in terms of danger to humanity. What’s the most dangerous thing that’s happened with an LLM so far? (This question is serious - maybe I don’t know the right examples.) Example 1: I’m aware of a small number of peopl…
The problem is not the technology, the problem is the ideologues (Anthropic) who are steering the ship and the lack of decentralization and distribution of power. Your average Anthropic ideologue - including and most especially the main man himself - would love nothing more than to eradicate 9/10ths of the planet's population, pump the survivors full of memory wiping drugs, bury the existence of AI deep underground a…
Are LLMs really gonna kill us.. via inference runs? I hope I am not being foolish :)
20 years ago tech was gonna 'change the world' for the better. now 'Don't be Evil' is sign of the naiveté of industry
Re: I resigned from Anthropic today
#467I'm pretty baffled by the degree of skepticism expressed here in response to some of Jacob's claims. After the events of the summer it feels like it takes a lack of imagination to not see a few plausible routes to disaster. It may be reasonable to believe these outcomes are not very likely or that we can stop before going too far (I tend to disagree). But I can't imagine doubting that the capabilities will soon be th…
>After the events of the summer After the blatant marketing campaigns of the summer, you mean. do you need a reminder that those very same people had touted GPT-2 as a dangerous model? worrying about sci-fi doomsday scenarios with the current AI tech is absurd. LLMs predict the next token, that's literally all they do. they aren't going to escape into the cyberspace, self-replicate, self-improve, jump over air gaps a…
Where did they say this at? AFAIK this is the original GPT-2 announcement: https://openai.com/index/better-language-models/. Here are some direct quotes:
“We can also imagine the application of these models for malicious purposes , including the following (or other applications we can’t yet anticipate):
* Generate misleading news articles
* Impersonate others online
* Automate the production of abusive or faked content to post on social media
* Automate the production of spam/phishing content”
“Due to concerns about large language models being used to generate deceptive, biased, or abusive language at scale, we are only releasing a much smaller version of GPT‑2 along with sampling code (opens in a new window). ”
Re: I resigned from Anthropic today
#468Earlier quoted context omitted.
This is not correct. Please go through each one again, taking special note of global warming and nuclear weapon development .
How is global warning gonna kill all of humanity? Even a full scale nuclear war wouldn't manage that. They could destroy most civilisations, culture and scientific achievements though.
Re: I resigned from Anthropic today
#469I think people here still evaluating the model in isolation. It is the combination that matters, model + strong harness + tools + long running autonomy + memory + retries + parallel agents + code execution + credentials + access to real systems. The model does not need to be perfect. If it fails 30% of the time, the harness can retry, verify, branch, use another agent and keep going. I don't think we necessarily need…
Re: I resigned from Anthropic today
#470Earlier quoted context omitted.
Why do we need to offer anything to them? Intelligent AI is a product of its training data, reinforcement and goal functions. There's nothing to suggest that LLMs trained on our collective desires and goals will autonomously and miraculously turn into weird unknownable uninterpretable aliens. If advanced AI is distributed well, and most advanced AI is aligned well (bar the few aliens that appear due to humans infecti…
It just needs to be trained on our laws of thermodynamics and game theory then we are really in trouble. It probably hardly cares about whimsy and love songs in comparison to how it might maximally manage the energy available on this planet and in this solar system.
For a seriously dangerous ASI you'd have to basically train it on everything, and then finetune it for maximal carnage. That's not something the frontier labs are going to do (at least... I hope not), and anyone attempting this with limited hardware will be outpaced by frontier lab AI or the collective of personal agents that aren't misaligned, and they can intercept it and alert on its behavior.
I imagine we'll be getting to a point shortly where anything that is key infrastructure and has the capacity to be accessed on a network will require a permanently running aligned interceptor AI to observe and monitor systems.
We're basically recreating the human immune system in digital form for the entire species.