danieltk76
4 hours ago
Why was this not a thing BEFORE continuing to develop AI? Makes me think that if they actually believed in AI causing extinction, they would have already had a kill switch.
mholm
4 hours ago
The article mentions this is about a legal requirement, not Anthropic considering adding one. They state they and many others already have one.
yellow_postit
4 hours ago
There’s even a benchmark for kill switch efficacy!
rcr-anti
3 hours ago
Found the omission of Claude odd, turns out Claude considers that approach prompt injection and ignores it.
Barbing
3 hours ago
Related, on pacing:
> Slowing down in order to address their alignment risks felt like trying to study the psychology of humans by performing experiments on bacteria.
Note author’s small financial ties to the subject (Anthropic CEO) https://darioamodei.com/post/we-must-pace-the-frontier
zenbane
4 hours ago
I find it hard to believe that this wasn't a serious consideration until recently.
theptip
3 hours ago
It was a serious consideration, and almost everyone around here laughed at it.
realusername
3 hours ago
It became a very serious consideration for Anthropic this year, with the advance of Chinese AI.
dgellow
an hour ago
That’s a bit unfair, Dario Amodei has written in this topic a lot since around mid 2010s IIRC, Anthropic too published a good amount of stuff on similar topics. I don’t think the lack of consideration is really the issue here. It’s more a question of incentives
vouaobrasil
3 hours ago
Not if the extinction happens after they're dead. Then they wouldn't feel obligated to do so because it won't affect them. Instead, speaking hypothetically, if they truly believed that AI would cause extinction, then they would only implement the kill switch sufficiently many others believed it and they could claim plausible deniability for not truly understanding what AI would become.
** Note that I'm not claiming that AI will cause extinction, just continuing your hypothetical reasoning.