Chris Olah 谈神经网络可解释性与 AI 安全

Chris Olah on Neural Network Interpretability and AI Safety

克里斯·奥拉 Chris Olah · 80,000 小时 · 2023-10-31 · 约 189 分钟 · 原视频 ↗

打开互动全文版(中英对照 + 朗读 + 问答)→

本期速览 · Overview

顶尖机器学习研究员 Chris Olah 探讨可解释性研究、神经网络工作原理、多模态神经元、缩放定律以及他的新 AI 实验室 Anthropic。

Chris Olah, a top machine learning researcher, discusses interpretability research, how neural networks work, multimodal neurons, scaling laws, and his new AI lab Anthropic.

要点 · TL;DR

核心观点 · Key points

反共识 · Contrarian takes

本期章节 · Chapters(共 69)

阅读全文双语转录 →