Share this page
-
DoomBench assesses “Chris Olah explains Anthropic's integrated large-model safety strategy” as evidence moving away from doom, with magnitude 30 and confidence 82 out of 100 in the safety and alignment category.
-
The DoomBench assessment of “Chris Olah explains Anthropic's integrated large-model safety strategy” is based on reporting from 80,000 Hours and records the editorial rationale, source quality, attribution, and revision history.
-
DoomBench summarizes “Chris Olah explains Anthropic's integrated large-model safety strategy” as follows: In a full interview, Chris Olah said large models were the greatest foreseeable source of AI risk and argued that safety research...
https://www.doombench.com/news/chris-olah-explains-anthropic-s-integrated-large-model-safety-strategy-2021-08-04