Loading jobs…
Loading jobs…
Anthropic
About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems.
About The Role
As an Enforcement Analyst focused on Radiological & Nuclear Harms, you will play a critical role in protecting against the misuse of AI systems for radiological and nuclear harms. You will enforce our Usage Policy with a specific focus on detecting and mitigating these risks, investigating potential violations, and help continuously strengthen our safeguards. The work sits at the intersection of radiological and nuclear threat analysis and platform enforcement: you will read real model interactions and make fast, well-reasoned calls about whether activity is benign research or a credible attempt at harm.
This role is a fit for someone who understands the dual-use nature of radiological and nuclear knowledge and enabling technologies well enough to separate the benign from the malicious — and who acts decisively under ambiguity. You will own and continuously improve the enforcement monitoring workflows for this harm area, and you will work closely with Policy, Threat Intelligence, Data Science, and Engineering cross-functional partners to accomplish this. Safety is core to our mission, and your work will directly protect individuals, communities, and critical systems from AI-facilitated weapons harm.
Important context for the role: In this position you may be exposed to and engage with explicit content spanning a range of topics, including material of a sexual, violent, or psychologically disturbing nature. The role also carries a shared on-call responsibility across the Policy and Enforcement teams.
Key Responsibilities
Enforce Usage Policies with a specific focus on detecting and mitigating potential radiological and nuclear risks. Take ownership of enforcement monitoring workflows for the radiological and nuclear harm area, improving end-to-end detection, investigation, triage, and escalation processes. Monitor and analyze platform activity to identify emerging patterns related to radiological and nuclear threats (within the broader CBRNE landscape) that may require policy updates or interventions.
Design and architect automated enforcement systems and review workflows that scale effectively while maintaining high accuracy across a technically complex content surface. Conduct thorough investigations of potential violations, gathering and documenting evidence to support enforcement decisions. Proactively surface trends and propose improvements to detection methods and review workflows.
Partner with Engineering and Data Science teams to optimize detection models and automated enforcement systems for policy violations. Partner with Policy and Threat Intelligence teams to understand potential exploits and contribute to risk-assessment frameworks, and partner with engineers iterating on safety systems. Provide enforcement-grounded feedback on policy gaps, and handle escalations and time-sensitive situations related to radiological and nuclear policy violations.
Minimum Qualifications
, physics, nuclear engineering, health physics, radiochemistry) and/or relevant professional experience in a related field. Possess experience in Trust & Safety, content moderation, or policy enforcement at platform scale, including working with generative AI tools to refine and optimize content review and enforcement workflows. Possess experience in utilizing AI tools to develop data dashboards for metrics collection in support of continuous improvement efforts.
Can analyze complex, ambiguous situations and make well-reasoned, defensible decisions under time pressure.