Kevin Roose
People are overreacting to the number here. Among AI lab employees, a p(doom) of 10% is fairly optimistic.
Evan Hubinger
Jacob is correct here—we really do earnestly believe AI could kill all humans! I personally think it is >10% within the next decade. I believe Anthropic is trying its best, but we do not yet have a plan to solve alignment for superintelligence and are not clearly on track to.