0D ZeroDay Careerssecurity roles · pay as posted · no accounts
Filters · AI security · safeguards & misuse · AI labs 45 open roles
clearance gone 45 open roles
45 open rolesAI security & safety · safeguards & misusewhole familyAlert meRSS
Engineering Manager, Safeguards
Anthropic ·London, UK · On-Site onsite London
£325k–£390k≈$423k–$507kPosted by the company in GBP. The US$ figure is approximate (rates as of 2026-09). Typical band for AI security & safety worldwide roles: $266k–$330k, from 82 postings. This one sits 56% above its midpoint. A rough placement, not a rating.
Safeguards Enforcement Lead, User Well-Being
Anthropic ·New York City, NY; Remote-Friendly (Travel-Requ… onsiteremote New YorkSan FranciscoWashington
~$245k–$285kEstimate from 12 posted bands in US — not the company's offer. Based on 12 similar postings, high confidence.
Safeguards Enforcement Analyst, Conventional Weapons
Anthropic ·New York City, NY; Remote-Friendly (Travel-Requ… onsiteremote New YorkSan FranciscoWashington
$245k–$330kPosted by the company. Typical band for AI security & safety · mid in US and US & Canada roles: $285k–$385k, from 59 postings. This one sits 14% below its midpoint. A rough placement, not a rating.
Cyber Operations Lead, Critical Harm Operations
OpenAI ·San Francisco remote San Francisco
$252k–$335kPosted by the company. Typical band for AI security & safety · staff in US and US & Canada roles: $178k–$251k, from 14 postings. This one sits 37% above its midpoint. A rough placement, not a rating.
Safeguards Enforcement Lead, Cyber Harms
Anthropic ·Remote-Friendly (Travel-Required) | Washington,… onsiteremote WashingtonSan FranciscoNew York
~$245k–$285kEstimate from 12 posted bands in US — not the company's offer. Based on 12 similar postings, high confidence.
Cyber Operations Strategist, Critical Harm Operations
OpenAI ·Ontario - Remote remote
CA$140k–CA$188k≈$102k–$137kPosted by the company in CAD. The US$ figure is approximate (rates as of 2026-09). Typical band for AI security & safety · mid in US & Canada roles: $285k–$385k, from 59 postings. This one sits 64% below its midpoint. A rough placement, not a rating.
Senior Safeguards Policy Lead, Cyber Harms
Anthropic ·Washington, DC · On-Site onsite Washington
~$295k–$383kEstimate from 12 posted bands in US — not the company's offer. Based on 12 similar postings, high confidence.
Safeguards Policy Analyst, Cyber Harms
Anthropic ·San Francisco, CA | Washington, DC · On-Site onsite San FranciscoWashington
~$245k–$285kEstimate from 12 posted bands in San Francisco · mid — not the company's offer. Based on 12 similar postings, high confidence.
Model Policy Manager
OpenAI ·San Francisco remote San Francisco
$207k–$295kPosted by the company. Typical band for AI security & safety · mid in US and US & Canada roles: $285k–$385k, from 59 postings. This one sits 25% below its midpoint. A rough placement, not a rating.
Technical Program Manager, AI Safety & Safeguards
OpenAI ·San Francisco onsite San Francisco
$257k–$445kPosted by the company. Typical band for AI security & safety · mid in US and US & Canada roles: $285k–$385k, from 59 postings. This one sits in line with it. A rough placement, not a rating.
Staff+ Software Engineer, Safeguards Data
Anthropic ·San Francisco, CA | New York City, NY · On-Site onsite San FranciscoNew York
~$305k–$395kEstimate from 12 posted bands in US — not the company's offer. Based on 12 similar postings, high confidence.
Staff+ Site Reliability Engineer, Safeguards ML Infra
Anthropic ·Remote-Friendly (Travel-Required) | San Francis… onsiteremote San FranciscoSeattleNew York
~$313k–$395kEstimate from 12 posted bands in US — not the company's offer. Based on 12 similar postings, high confidence.
Program Manager, Critical Harm Operations
OpenAI ·San Francisco remote San Francisco
$252k–$280kPosted by the company. Typical band for AI security & safety · mid in US and US & Canada roles: $285k–$385k, from 59 postings. This one sits 21% below its midpoint. A rough placement, not a rating.
Safeguards Enforcement Analyst, User Well-being
Anthropic ·Remote-Friendly, United States; San Francisco, … onsiteremote San FranciscoNew YorkWashington
$245k–$285kPosted by the company. Typical band for AI security & safety · mid in US and US & Canada roles: $285k–$385k, from 59 postings. This one sits 21% below its midpoint. A rough placement, not a rating.
Machine Learning Infrastructure Engineer, Safeguards Research
Anthropic ·San Francisco, CA | New York City, NY · San Fra… onsite San FranciscoNew York
$350k–$500kPosted by the company. Typical band for AI security & safety · mid in US and US & Canada roles: $285k–$385k, from 59 postings. This one sits 27% above its midpoint. A rough placement, not a rating.
Safeguards Enforcement Analyst, Violence & Extremism
Anthropic ·Remote-Friendly, United States; San Francisco, … onsiteremote San FranciscoNew YorkWashington
$285k–$330kPosted by the company. Typical band for AI security & safety · mid in US and US & Canada roles: $285k–$385k, from 59 postings. This one sits in line with it. A rough placement, not a rating.
Safeguards Enforcement Analyst, Fraud & Scams
Anthropic ·Remote-Friendly, United States; San Francisco, … onsiteremote San FranciscoNew YorkWashington
$245k–$285kPosted by the company. Typical band for AI security & safety · mid in US and US & Canada roles: $285k–$385k, from 59 postings. This one sits 21% below its midpoint. A rough placement, not a rating.
Safeguards Enforcement Analyst, Ban Evasion & Recidivism
Anthropic ·Remote-Friendly, United States; San Francisco, … onsiteremote San FranciscoNew YorkWashington
$245k–$285kPosted by the company. Typical band for AI security & safety · mid in US and US & Canada roles: $285k–$385k, from 59 postings. This one sits 21% below its midpoint. A rough placement, not a rating.
Safeguards Enforcement Analyst, Account Takeover & Credential Abuse
Anthropic ·Remote-Friendly, United States; San Francisco, … onsiteremote San FranciscoNew YorkWashington
$245k–$285kPosted by the company. Typical band for AI security & safety · mid in US and US & Canada roles: $285k–$385k, from 59 postings. This one sits 21% below its midpoint. A rough placement, not a rating.
Safeguards Enforcement Analyst, Access Controls & Identity
Anthropic ·Remote-Friendly, United States; San Francisco, … onsiteremote San FranciscoNew YorkWashington
$285k–$330kPosted by the company. Typical band for AI security & safety · mid in US and US & Canada roles: $285k–$385k, from 59 postings. This one sits in line with it. A rough placement, not a rating.
Staff+ Software Engineer, Safeguards Review Tooling
Anthropic ·San Francisco, CA · On-Site onsite San Francisco
~$305k–$395kEstimate from 12 posted bands in US — not the company's offer. Based on 12 similar postings, high confidence.
Safeguards Enforcement Analyst, Radiological & Nuclear Harms
Anthropic ·Remote-Friendly, United States; San Francisco, … onsiteremote San FranciscoNew YorkWashington
$245k–$285kPosted by the company. Typical band for AI security & safety · mid in US and US & Canada roles: $285k–$385k, from 59 postings. This one sits 21% below its midpoint. A rough placement, not a rating.
Safeguards Enforcement Analyst, Bio Harms
Anthropic ·Remote-Friendly, United States; San Francisco, … onsiteremote San FranciscoNew YorkWashington
$245k–$285kPosted by the company. Typical band for AI security & safety · mid in US and US & Canada roles: $285k–$385k, from 59 postings. This one sits 21% below its midpoint. A rough placement, not a rating.
Safeguards Enforcement Analyst, Chem & Explosives Harms
Anthropic ·Remote-Friendly, United States; San Francisco, … onsiteremote San FranciscoNew YorkWashington
$245k–$285kPosted by the company. Typical band for AI security & safety · mid in US and US & Canada roles: $285k–$385k, from 59 postings. This one sits 21% below its midpoint. A rough placement, not a rating.
Safeguards Enforcement Analyst, Child Safety
Anthropic ·Remote-Friendly, United States; San Francisco, … onsiteremote San FranciscoNew YorkWashington
$245k–$285kPosted by the company. Typical band for AI security & safety · mid in US and US & Canada roles: $285k–$385k, from 59 postings. This one sits 21% below its midpoint. A rough placement, not a rating.
Safeguards Enforcement Analyst, Age-Appropriate Design
Anthropic ·Remote-Friendly, United States; San Francisco, … onsiteremote San FranciscoNew YorkWashington
$245k–$285kPosted by the company. Typical band for AI security & safety · mid in US and US & Canada roles: $285k–$385k, from 59 postings. This one sits 21% below its midpoint. A rough placement, not a rating.
Safeguards Enforcement Analyst, Cyber Harm
Anthropic ·Remote-Friendly, United States; San Francisco, … onsiteremote San FranciscoNew YorkWashington
$285k–$330kPosted by the company. Typical band for AI security & safety · mid in US and US & Canada roles: $285k–$385k, from 59 postings. This one sits in line with it. A rough placement, not a rating.
Safeguards Enforcement Analyst, Integrity & Authenticity
Anthropic ·Remote-Friendly, United States; San Francisco, … onsiteremote San FranciscoNew YorkWashington
$285k–$330kPosted by the company. Typical band for AI security & safety · mid in US and US & Canada roles: $285k–$385k, from 59 postings. This one sits in line with it. A rough placement, not a rating.
Red Team Engineer, Safeguards
Anthropic ·Remote-Friendly (Travel Required) | San Francis… onsiteremote San Francisco
$320k–$405kPosted by the company. Typical band for AI security & safety · mid in US and US & Canada roles: $285k–$385k, from 59 postings. This one sits in line with it. A rough placement, not a rating.
Product Manager, Safeguards (Child Safety)
Anthropic ·San Francisco, CA · On-Site onsite San Francisco
$305k–$385kPosted by the company. Typical band for AI security & safety · mid in US and US & Canada roles: $285k–$385k, from 59 postings. This one sits in line with it. A rough placement, not a rating.
Data Engineer, Safeguards
Anthropic ·San Francisco, CA | New York City, NY · San Fra… onsite San FranciscoNew York
$320k–$405kPosted by the company. Typical band for AI security & safety · mid in US and US & Canada roles: $285k–$385k, from 59 postings. This one sits in line with it. A rough placement, not a rating.
Product Manager, Safeguards Rare Harms
Anthropic ·San Francisco, CA · On-Site onsite San Francisco
$305k–$385kPosted by the company. Typical band for AI security & safety · mid in US and US & Canada roles: $285k–$385k, from 59 postings. This one sits in line with it. A rough placement, not a rating.
Engineering Manager, Safeguards Review Tooling
Anthropic ·San Francisco, CA onsite San Francisco
$405k–$485kPosted by the company. Typical band for AI security & safety in US and US & Canada roles: $266k–$330k, from 78 postings. This one sits 49% above its midpoint. A rough placement, not a rating.
Software Engineer, Scaled Abuse
OpenAI ·San Francisco onsite San Francisco
$266k–$445kPosted by the company. Typical band for AI security & safety · mid in US and US & Canada roles: $285k–$385k, from 59 postings. This one sits in line with it. A rough placement, not a rating.
Data Scientist, Safeguards
Anthropic ·New York City, NY; San Francisco, CA; Seattle, … onsiteremote New YorkSan FranciscoSeattle
$285k–$380kPosted by the company. Typical band for AI security & safety · mid in US and US & Canada roles: $285k–$385k, from 59 postings. This one sits in line with it. A rough placement, not a rating.
Model Policy, Frontier Cyber Risk
OpenAI ·San Francisco remote San Francisco
$266k–$335kPosted by the company. Typical band for AI security & safety · mid in US and US & Canada roles: $285k–$385k, from 59 postings. This one sits 10% below its midpoint. A rough placement, not a rating.
Safeguards Enforcement Analyst, Safety Evaluations
Anthropic ·Remote-Friendly (Travel-Required) | San Francis… onsiteremote San FranciscoWashingtonNew York
$230k–$270kPosted by the company. Typical band for AI security & safety · mid in US and US & Canada roles: $285k–$385k, from 59 postings. This one sits 25% below its midpoint. A rough placement, not a rating.
Data Scientist, Preparedness
OpenAI ·San Francisco onsite San Francisco
$345k–$385kPosted by the company. Typical band for AI security & safety · mid in US and US & Canada roles: $285k–$385k, from 59 postings. This one sits in line with it. A rough placement, not a rating.
Model Policy, Chemical & Biological Risk
OpenAI ·San Francisco remote San Francisco
$207k–$295kPosted by the company. Typical band for AI security & safety · mid in US and US & Canada roles: $285k–$385k, from 59 postings. This one sits 25% below its midpoint. A rough placement, not a rating.
Technical Program Manager, Safeguards (Infrastructure & Evals)
Anthropic ·San Francisco, CA | New York City, NY | Seattle… onsite San FranciscoNew YorkSeattle
~$305k–$385kEstimate from 12 posted bands in San Francisco · mid — not the company's offer. Based on 12 similar postings, high confidence.
Product Manager, Safeguards (Cyber)
Anthropic ·San Francisco, CA · On-Site onsite San Francisco
$305k–$385kPosted by the company. Typical band for AI security & safety · mid in US and US & Canada roles: $285k–$385k, from 59 postings. This one sits in line with it. A rough placement, not a rating.
Staff+ Software Engineer, Safeguards Infrastructure
Anthropic ·London, UK · On-Site onsite London
not postedToo few comparable posted roles for an estimate
Staff+ Software Engineer, Safeguards
Anthropic ·San Francisco, CA | New York City, NY · On-Site onsite San FranciscoNew York
~$305k–$395kEstimate from 12 posted bands in US — not the company's offer. Based on 12 similar postings, high confidence.
ML/Research Engineer, Safeguards
Anthropic ·San Francisco, CA | New York City, NY · San Fra… onsite San FranciscoNew York
$350k–$500kPosted by the company. Typical band for AI security & safety · mid in US and US & Canada roles: $285k–$385k, from 59 postings. This one sits 27% above its midpoint. A rough placement, not a rating.
Staff+ Software Engineer, Safeguards ML Infrastructure
Anthropic ·San Francisco, CA · On-Site onsite San Francisco
~$305k–$395kEstimate from 12 posted bands in US — not the company's offer. Based on 12 similar postings, high confidence.