Staff+ Software Engineer, Safeguards Review Tooling na Anthropic

Híbrido - San Francisco, CA

Candidatar-se
Ver mais vagas na Anthropic

Anthropic is looking for a Staff+ Software Engineer for the Safeguards Review Tooling team. In this foundational role, you will own the tools our safety investigators rely on to understand platform behavior and act on potential harms. You’ll build and scale investigation, review, and enforcement tooling, develop reusable APIs and backend services, partner with cross‑functional stakeholders, and instrument tooling to surface key metrics—all while ensuring the system adheres to privacy, compliance, and operational requirements.

Salary

USD 320,000 - 485,000

Requirements

Skills

  • A technical background in full-stack or platform engineering, with the ability to engage deeply in architecture and design discussions
  • Experience shipping internal tools or platforms with demanding operational users, and a track record of improving their workflows measurably
  • Experience working cross‑functionally with non-engineering partners such as operations, policy, or legal teams
  • Excellent communication skills, including the ability to explain technical tradeoffs to non-technical stakeholders
  • Care about the societal impacts of AI and want your work to make powerful systems safer
  • 8+ years of industry software engineering experience
  • Experience building trust and safety, integrity, fraud, or abuse‑prevention tooling, or other systems supporting human review at scale
  • Experience designing systems under strict privacy, compliance, or data governance constraints, such as zero data retention environments
  • Experience integrating LLMs or agentic systems into operational workflows, or building human-in-the-loop automation—including using agentic coding tools (e.g., Claude Code) as a core part of your own development workflow
  • Experience building developer platforms or extensible tooling frameworks that other teams build on top of
  • Experience supporting enforcement or moderation systems across multiple product surfaces, including enterprise or cloud platform contexts
  • A product‑minded approach to internal users: you work directly with the people using your tools, watch where they struggle, and fix it

Responsibilities

  • Build investigation, review, and enforcement tooling for both first-party and third-party platform surfaces—including case queues, investigation views, decision and audit logging, and account-actioning workflows
  • Develop the platform layer of reusable APIs, data storage, and backend services that let new review workflows be stood up quickly and safely
  • Scale review through automation, including enabling reviewers to use Claude effectively and building toward Claude-assisted and Claude-driven review workflows
  • Partner with policy, operations, legal, privacy, and data science stakeholders to translate enforcement and investigation needs into reliable, well‑designed systems that measurably reduce handling time and decision error
  • Build in the guardrails that sensitive internal tools require: granular permissions, audit trails, data‑access controls, and reviewer wellbeing features such as content obfuscation and exposure limits
  • Instrument the tools you ship—surfacing metrics on queue health, reviewer throughput, and decision quality—and ensure tooling evolves alongside new privacy primitives and data retention commitments

Technologies

ClaudeLLMClaude Code

Compartilhar vaga

Descubra se seu currículo está pronto para esta vaga

Veja como nossa IA pode otimizar seu currículo e aumentar suas chances de conseguir esta posição.