Updated · 1 episodes · 1 show · 1 source notes
Evan Hubinger
Overview
Evan Hubinger appears in this wiki through the All-In episode as an Anthropic alignment-safety figure whose public comments made the Coxon controversy more significant for the hosts.
Current Profile
The source says Hubinger leads Alignment Science at Anthropic and publicly stated that AI could kill all humans, putting his own estimate above 10% within the next decade. The hosts treat his endorsement as more legally and commercially important than Coxon’s because he remained inside Anthropic and therefore sharpened the company’s Frontier AI IPO Disclosure Risk.
Key Characteristics
- Anthropic alignment-safety leader in the source’s account.
- Publicly validates a nontrivial AI extinction-risk estimate.
- Makes the controversy harder for Anthropic to dismiss as only a short-tenured former employee’s claim.
- Connects Anthropic’s safety culture to IPO, disclosure, and liability questions.
Evidence
- Role and claim: AI Kills Everybody or Doomer Psyop? OpenAI’s Math Breakthrough, Nike’s $200B Collapse says Hubinger leads Alignment Science at Anthropic and estimated more than 10% AI-kills-all risk within a decade.
- Disclosure relevance: AI Kills Everybody or Doomer Psyop? OpenAI’s Math Breakthrough, Nike’s $200B Collapse records Sacks arguing Hubinger’s support creates greater legal significance than Coxon’s post alone.
- Culture conflict: AI Kills Everybody or Doomer Psyop? OpenAI’s Math Breakthrough, Nike’s $200B Collapse says Anthropic may have difficulty rejecting Coxon-like claims if many safety employees agree.
Qualifications
This page is based on a single host-mediated source and does not independently verify Hubinger’s exact title, statement, probability estimate, or legal implications.
What Changed
- Created the entity from the All-In Anthropic disclosure-risk discussion.
Relationships
- Anthropic - employer and organizational context in the source.
- Jacob Coxon - resignation-thread figure whose claims Hubinger is said to have validated.
- AI Doomerism - risk narrative connected to Hubinger’s public comments.
- Frontier AI IPO Disclosure Risk - capital-markets implication attached to his role.
- AI Safety Narrative Backfire - broader risk when safety claims damage trust or invite control.
Sources
1 source notes across 1 show
- AI Kills Everybody or Doomer Psyop? OpenAI's Math Breakthrough, Nike's $200B Collapse All-In with Chamath, Jason, Sacks & Friedberg