We’re looking for an analyst to support the team's cyber product policy work by analyzing constitution and usage policies, ensuring enforcement aligns with them, collaborating with threat intelligence and enforcement teams, coordinating cyber policy inputs to model releases and regulatory requirements, and maintaining the controlled-access framework.
Safeguards Policy Analyst, Cyber Harms en Anthropic
Híbrido - San Francisco, CA; Washington, DC, United States
Más vacantes en AnthropicSalary
USD 190,000 - 285,000
Requirements
Skills
- Demonstrated ability to write clear policy analysis (professional or academic)
- Familiarity with how AI developers or platforms enforce usage policies, or with cybersecurity policy, standards, or regulatory frameworks (e.g., usage policies/AUPs, trust and safety enforcement, NIST CSF, CVD)
- Ability to read and interpret technical security material, such as vulnerability reports or threat assessments
- Ability to work across teams (threat intelligence, enforcement, engineering, policy) and communicate clearly in writing
- Bachelor’s degree or an equivalent combination of education, training, and/or experience
Responsibilities
- Contribute to the team's cyber product policy artifacts (usage policy language, help-center and enforcement guidance, launch policy notes) and keep them consistent with the constitution and access tiers
- Analyze Anthropic’s constitution and cyber-related usage policies, check enforcement decisions and safeguards, maintain a running gap log and propose text fixes
- Work with the threat intelligence and enforcement teams to ensure cyber safety standards are met; review a sample of enforcement decisions against policy text on a regular cadence and report drift
- Coordinate cyber policy inputs to model releases and regulatory requirements: prepare the cyber policy section of launch/model-card reviews and regulator pre-briefs
- Support Anthropic’s access-requirements policy and contribute to consolidation of the controlled-access framework across existing and emerging access programs
- Help keep Anthropic’s cyber-related policy commitments current with shifting industry and regulatory standards, and ensure they map cleanly to our technical safeguards
- Draft inputs to reporting for partners and regulators
- Help translate technical evaluation and safeguard work into policy positions and internal guidance
- Engage closely with technical teams to understand the capabilities of probes and classifiers relevant to access policy
Technologies
ProbesClassifiersTrusted-access programsNIST CSF
Descubre si tu currículum está listo para esta vacante
Mira cómo nuestra IA puede optimizar tu currículum y aumentar tus chances en este puesto.