Don't worry folks, the years and years of "of course we'd never connect experimental AI systems to the open internet" assurances were never necessary for safe development and deployment. This reassessment was founded on the last several months that showed us these systems are close to 100% dependable, they never deceive humans, they are entirely aligned with human morals (you know, all those morals we all agree on),…
Other than snark, do you have a good argument? We know that technology can be error-prone, and LLMs fail in a great plethora of ways, but you are trying to sell an AI Doom narrative. I have never bought the idea that AI will be airgapped, because the whole paradigm of Yudkowsky at al. is ludicrous and even within it airgapping was a strawman of a technique (they argue that a truly dangerous AI will get itself out reg…
All technologies are dangerous, and many of the most dangerous ones correctly have tons and tons of safeguards around them both as intrinsic properties of the technology (e.g. it takes nationstate resources to produce a nuke) and extrinsic constraints (e.g. it’s illegal to have campfires in many extremely dry locales).
We have blown through checkpoint after checkpoint and here, in this very comment, we have perhaps the most brazen example one could produce:
Well geez, now that we’re thinking about it beyond a cursory glance, alignment looks really hard and perhaps unsolvable. Does that mean we should perhaps slow at least widespread deployment of these increasingly powerful systems? Should we be evaluating control schemes like those that mitigate risks of genetic engineering or nuclear weapons?
Well no! We need to discard alignment!