Safeguards Policy Analyst, Cyber Harms at Anthropic

Hybrid - San Francisco, CA; Washington, DC, United States

Apply
More jobs at Anthropic

We’re looking for an analyst to support the team's cyber product policy work by analyzing constitution and usage policies, ensuring enforcement aligns with them, collaborating with threat intelligence and enforcement teams, coordinating cyber policy inputs to model releases and regulatory requirements, and maintaining the controlled-access framework.

Salary

USD 190,000 - 285,000

Requirements

Skills

  • Demonstrated ability to write clear policy analysis (professional or academic)
  • Familiarity with how AI developers or platforms enforce usage policies, or with cybersecurity policy, standards, or regulatory frameworks (e.g., usage policies/AUPs, trust and safety enforcement, NIST CSF, CVD)
  • Ability to read and interpret technical security material, such as vulnerability reports or threat assessments
  • Ability to work across teams (threat intelligence, enforcement, engineering, policy) and communicate clearly in writing
  • Bachelor’s degree or an equivalent combination of education, training, and/or experience

Responsibilities

  • Contribute to the team's cyber product policy artifacts (usage policy language, help-center and enforcement guidance, launch policy notes) and keep them consistent with the constitution and access tiers
  • Analyze Anthropic’s constitution and cyber-related usage policies, check enforcement decisions and safeguards, maintain a running gap log and propose text fixes
  • Work with the threat intelligence and enforcement teams to ensure cyber safety standards are met; review a sample of enforcement decisions against policy text on a regular cadence and report drift
  • Coordinate cyber policy inputs to model releases and regulatory requirements: prepare the cyber policy section of launch/model-card reviews and regulator pre-briefs
  • Support Anthropic’s access-requirements policy and contribute to consolidation of the controlled-access framework across existing and emerging access programs
  • Help keep Anthropic’s cyber-related policy commitments current with shifting industry and regulatory standards, and ensure they map cleanly to our technical safeguards
  • Draft inputs to reporting for partners and regulators
  • Help translate technical evaluation and safeguard work into policy positions and internal guidance
  • Engage closely with technical teams to understand the capabilities of probes and classifiers relevant to access policy

Technologies

ProbesClassifiersTrusted-access programsNIST CSF

See if your resume is ready for this job

See how our AI can optimize your resume and improve your chances for this role.