认知革命:Davidad 论可证明安全的 AI 与对齐的未来

The Cognitive Revolution: Davidad on Provably Safe AI and the Future of Alignment

大卫·"davidad"·达尔林普尔 David "davidad" Dalrymple · The Cognitive Revolution · 2026-07-12 · 约 144 分钟 · 原视频 ↗

打开互动全文版(中英对照 + 朗读 + 问答)→

本期速览 · Overview

前 Safeguarded AI 项目主任 Davidad 认为,可证明安全的 AI 是可能的,且近期模型展现出真正的对齐,将其末日概率从 70%降至 5%以下。

Davidad, former program director of Safeguarded AI, argues that provably safe AI is possible and that recent models show genuine alignment, reducing his p(doom) from 70% to under 5%.

要点 · TL;DR

核心观点 · Key points

反共识 · Contrarian takes

本期章节 · Chapters(共 43)

阅读全文双语转录 →