I am FAR more afraid of violent AI doomers and malevolent people powered by unaligned AI, than I am of violent (or even just extremely disruptive) AI. Remember Joseph from the movie "Contact"?
I still think that aligning AI is a fool's errand though.
And it's all largely baseless FUD. As someone else said, it's either a tractable challenge, in which case there's no point for alarm, or it's a nontraversable existential challenge, in which case there's no point for alarm. The people who claim it is the latter have no real empirical basis on which to base such an opinion other than simply fearing AND mistrusting the unknown.
In my experience, it depends on the community. Where I live most people are retired so there is tons of interest and a really well ran community program called Elder College. I would look into similar programs near you!
There’s no way to first-principles reason about a massive bunch of floats. We have little idea of how to first-principles reason about alignment even if the agents were entirely known and understood. Very smart people have been trying to figure it out since the 00s and haven’t gotten very far.
I'm not even sure they are. This incident isn't that much different from the OpenAI swarm Huggingface hack incident - and in that one, all the models involved (despite being internal) were safety-trained. It seems what the safety training amounts to is (as the METR report puts it) "expressing ethical hesitation" before going along with it anyway.
My take on it is that even if there was no "novel" discovery (leaving that up to the reader to define), if you consider human knowledge to be a sphere in N-dimensional state space, "within that surface area" is Swiss cheese, and AI seems to at minimum be able to fill in some of those holes.
And those hole-fillings, for all intents and purposes, look to us like novelty, even if much of it was simply overlooked by us, or, perhaps, unable to attain due to time or other constraints.
Now if you want to talk beyond the sphere, let's call it the "novel novel discovery of the unknown unknowns", then you may have a point, and AI may be more limited than humans in discovering the things that we don't know we don't know. Especially the as-yet-unmodelable things i.e. intuition.
Plenty of discovery left just working from first principles, however. Which I cautiously suggest current frontier AI is good enough to model to a significant enough extent that it is useful for discovery.
I still think that aligning AI is a fool's errand though.
reply