Dan Hendrycks πΊπΈ
AI-safety researcher whose work concerns evaluations, societal-scale risk, robustness, and safeguards for advanced systems.
Center for AI Safety
- Evidence items
- 4
- Toward pressure
- +0.02
- Away pressure
- β0.05
- Net attributed pressure
- -0.04
0 comments Β· 0 votes
Sign in to join the discussion β
No comments yet. Start the discussion.
First-person and official sources
These sources guide discovery. A statement still needs a dated, attributable, source-backed evidence assessment before it can affect the index.
- institutional profileCenter for AI Safety
Assessments involving Dan Hendrycks
Hendrycks proposes deterrence and nonproliferation for superintelligence
Hendrycks, Eric Schmidt, and Alexandr Wang propose a strategy combining deterrence against destabilizing AI projects, controls on chips and model weights, technical misuse safeguards, legal rules for agents, and measures to limit automation-driven instability.
xAI appoints Dan Hendrycks as an external safety adviser
xAI named Center for AI Safety director Dan Hendrycks as an adviser when it disclosed its founding team, adding an external safety specialist to a new frontier laboratory entering an already competitive model race.
Hendrycks argues competition may select for less controllable AI
Hendrycks argues that commercial and national competition can reward increasingly autonomous systems, weaken oversight, select against systems that are easy to stop, and create dependence on AI embedded in critical functions.
Hendrycks and Mazeika publish a structured AI x-risk analysis method
The paper adapts established hazard analysis and systems-safety concepts to advanced AI, proposes X-Risk Sheets for assessing safety research, and warns that safety work can backfire when it improves general capabilities more than control.