Skip to main content
ResumeKart
← Back to Jobs

Staff+ Software Engineer, Safeguards Review Tooling

Anthropic•San Francisco, CA
Full-timeLead
👁️ 0 views•📝 0 applications•Posted 7/14/2026•Expires 10/13/2026
Tailor Resume for This JobCheck ATS Score

Get alerts for roles like this

More Staff+ Software Engineer, Safeguards Review Tooling roles in San Francisco, CA — straight to your inbox. No account needed.

Applying to this role? Tailor your résumé to this job description in one click, then download it clean — no watermark, no subscription.

Job Description

About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems. About the role The Safeguards team is responsible for ensuring our models and products are developed and deployed safely. We're looking for engineers for our Review Tooling team, which builds the systems that humans — and increasingly Claude — use to investigate potential harms and take enforcement actions across Anthropic's first-party products and third-party cloud platforms. This is a foundational role: as one of the first engineers on this new team, you'll own the tools our safety investigators rely on to understand what's happening on our platforms and act on it, as well as the platform underneath those tools. That platform includes analytics capabilities, privacy-preserving primitives that keep review workflows compatible with our data retention commitments, and a sandbox environment where new review interfaces and workflows can be built and iterated quickly. As model capabilities and usage grow, you'll also drive how we scale review through automation — building systems where Claude meaningfully extends what human reviewers can do, while keeping people in the loop where their judgment matters most. These are internal tools, but they are anything but low-stakes: the speed, clarity, and reliability of this tooling directly determines how quickly Anthropic can identify harmful behavior, make sound enforcement decisions, and feed signal back into model training. You'll partner closely with policy, operations, data science, legal, and privacy teams to ensure our enforcement systems are effective, accurate, and trustworthy. Key responsi

Required Skills

AuditCommunicationData ScienceLogisticsLogistics ManagementRecruitment

Partner picks for Staff+ Software Engineer, Safeguards Review Tooling in San Francisco

Matched to the skills this page calls for and the candidate's location.

Partner
  • edXVerified partner
    Partner course provider

    Courses and programmes from universities and institutions worldwide.

    covers data science
  • Partner course provider

    Free online courses from IITs and IISc, with proctored certification exams.

    covers data science
  • Partner course provider

    Online higher-education programmes with university partners in India and abroad.

    covers data science

Partners are ResumeKart affiliates or institutes it works with; ResumeKart may earn a commission when a candidate enrols. Placement is decided by relevance, not payment. How ResumeKart earns

The best-paying roles in your field. Every week. Free.

Join 10,000+ professionals getting job alerts and salary insights in their inbox

We respect your privacy. Unsubscribe anytime with one click.