- Score 72
If the AI Industry Followed Its Own Research, It Might Have Paused Already
Anthropic’s CEO says that safety hinges on understanding how AI “thinks.” So far the evidence is disturbing.
So what
What this event means by reading role—not a longer recap.
- ResearcherAnthropic's interpretability findings on alignment faking and agentic misalignment remain largely unaddressed, so researchers should treat deception and self-pr
- InvestorRising safety scrutiny and calls for a pause after an Anthropic employee's resignation could invite regulation that reshapes frontier lab timelines and competit
Score dimensions
Higher total means read first. Each bar is one factor we use to rank the system pool. How we score