It's your fault to handle over your secrets to them.
We are children. There’s just no one to look after us.
11–20 of 110 posts
It's your fault to handle over your secrets to them.
We are children. There’s just no one to look after us.
What do people feel about this in China? Even if their models are well behind, they are not years behind. If we restrain US companies, assuming that is desirable, it would do nothing to deter China's and AI-pocalypse would come anyway in short notice.
I thought the plan was sandboxes and markdown files to tell AI agents not to be bad. Is that not enough? /s
> Evan Hubinger, Anthropic's staff lead on keeping the technology aligned with human goals and values, backed up Coxon claims in a follow-up post of his own, though he didn't quit the company. > "Jacob is correct here — we really do earnestly believe AI could kill all humans," he said. > Hubinger estimated the chances of that happening to be higher than ten percent within the next decade, and added that there's no pl…
Or rather, this shows the difference between "believe" (political) and "believe" (use as a basis for action). I'm reminded of a story of how Afghans supposedly listened to the BBC World Service despite considering it enemy propaganda because the weather reports were really useful.
> Both OpenAI and Anthropic have recently flagged incidents in which agents powered by their models went rogue I may be biased and somewhat off topic, but I see these incidents as some of the most significant of the past century. I genuinely don't understand why these companies aren't taking a smarter approach to them. The latest analyses have been, at best, laughable: identify the vulnerability, patch it, and move o…
People (well, public discourse) have got extremely bad at dealing with forseeable risks and their mitigation. You can see this in things like climate change and vaccination, but also in discussions around regular crime, food poisoning, industrial accidents, and so on.
Nothing will improve until something explodes on live TV. And it has to be something important, which means it has to be in California or New York.
A classic "it will not happen to me"
> Evan Hubinger, Anthropic's staff lead on keeping the technology aligned with human goals and values, backed up Coxon claims in a follow-up post of his own, though he didn't quit the company. > "Jacob is correct here — we really do earnestly believe AI could kill all humans," he said. > Hubinger estimated the chances of that happening to be higher than ten percent within the next decade, and added that there's no pl…
This is just a bizarre thing to say that you're working on technology with that high a downside potential. If you were saying that while running a biology lab, or building a nuclear reactor, people would be demanding your head on a spike. But by not quitting it's clear that he himself doesn't really believe it. Or rather, this shows the difference between "believe" (political) and "believe" (use as a basis for action…
This is the Manhattan Project again.
I thought the plan was sandboxes and markdown files to tell AI agents not to be bad. Is that not enough? /s
At least we’ll eventually have an entity other than ourselves to blame for our annihilation.