Safeguards Research Team (launched July 2026) — A dedicated new team focused on jailbreak robustness, automated red teaming, and monitoring techniques for both misuse and misalignment . Roles include Applied Safety Research Engineer, Safeguards, requiring 4+ years of software/ML engineering experience and Python proficiency .
Domain-specific Safeguards Enforcement Analysts — The most distinctive initiative:
In March 2026, Anthropic specifically recruited a chemical weapons and explosives expert to design guardrails, seeking applicants with at least five years of experience in chemical weapons and explosives defense .
Anthropic Fellows Program (2026 cohorts) — An open call for engineers and researchers to investigate Anthropic's highest-priority safety research questions, with funding and mentorship .
Trust & Safety Research Engineers — Building ML models to detect unwanted or anomalous user and API partner behaviors in production .
Across 264 public postings (2025–2026), the median compensation is $356,000 ($289,000–$403,000), with some technical staff earning over $1 million .
| Role | Base Pay Range | Key Qualifications |
|---|---|---|
| Safeguards Enforcement Analyst, Chem & Explosives | $245K–$285K | Domain expertise in chemical/explosives safety |
| Applied Safety Research Engineer, Safeguards | $320K–$405K | 4+ yrs SWE/ML engineering, Python |
| Research Engineer / Scientist, Safeguards | $315K–$560K | Advanced degree, safety research experience |
| Research Engineer, Trust & Safety | $300K–$450K | ML modeling, production systems experience |
| AI Safety Researcher (Senior) | $200K–$300K | Senior-level red-teaming/safety background |
All roles are in San Francisco (in-office or hybrid), with equity on top of base pay.
This hiring push is a direct operational response to Amodei's escalating public warnings:
The hiring of weapons experts, cybercrime analysts, and safeguards engineers is the practical manifestation of these warnings.
OpenAI's parallel moves: In December 2025, OpenAI posted a Head of Preparedness role with a base pay range up to $555,000 — and in February 2026, they filled it by poaching an Anthropic safety researcher . In March 2026, OpenAI also posted job openings for chemical weapons and explosives experts, closely mirroring Anthropic's domain-specific misuse roles . The two companies collaborated on a first-of-its-kind joint safety evaluation in mid-2025, where each tested the other's public models for misalignment propensities .
Key differences: Anthropic's approach is more structurally embedded — it created a permanent Safeguards Research Team and a fellows pipeline, rather than hiring individual senior leads. Bloomberg reported in March 2026 that the combined safety headcount at both companies remains strikingly small, with investment in safety-oriented roles "disturbingly" imbalanced relative to product-focused hiring . Anthropic frames its hiring as catastrophe prevention (explicitly listing weapons of mass destruction scenarios), whereas OpenAI has tended to emphasize responsible deployment and preparedness .
The broader self-regulation context: No comprehensive federal AI safety legislation has passed in the U.S. as of mid-2026. Both companies operate under voluntary commitments (the White House's 2023 voluntary pledges, Anthropic's published Transparency Hub commitments on red teaming, child safety, and whistleblowing policies) . Amodei has been publicly candid that self-regulation is insufficient — he told 60 Minutes he is "deeply uncomfortable" with C-suites making these decisions alone — yet in the absence of legislation, both firms are building out misuse prevention teams as a de facto private-sector safety net. The result is a race to hire scarce domain experts (chemical weapons specialists, nuclear analysts) that neither the government nor the private sector has in large supply, while the underlying talent pool remains very thin .