Jobs · Information Technology · California

Technical Program Manager, Safeguards (Infrastructure & Evals)

Anthropic · San Francisco, CA · 2 wk ago
HybridInformation Technology$290k–$365k/yrFull-time

About the role

Safeguards Engineering builds and operates the infrastructure that keeps Anthropic's AI systems safe in production. The role involves driving reliability, owning incident response and post-mortem processes, and coordinating platform investments.

Responsibilities

  • Own the Safeguards Engineering ops review - Drive the recurring cadence that keeps the team informed and coordinated.
  • Drive incident tracking and post-mortem execution - Ensure incidents are tracked and post-mortems are completed.
  • Establish and maintain Service Level Objectives (SLOs) with partner teams - Define and monitor SLOs for safety-critical pipelines.
  • Maintain runbook quality and incident-ownership clarity - Ensure runbooks are accurate and incidents are clearly assigned.
  • Drive platform migrations and infrastructure projects - Manage migrations and infrastructure improvements across teams.
  • Partner with the evals engineering team - Drive improvements to the evaluation platform and manage dependencies.

Requirements

Strong technical program management experience, particularly in operational or infrastructure-heavy environments. Understanding of how production ML systems work and ability to triage incidents intelligently. Energized by closing loops and effective coordination across team boundaries.

Qualifications

  • Solid technical program management experience.
  • Understanding of how production ML systems work.
  • Experience with or strong interest in AI safety.
  • Experience with SRE practices, incident management frameworks, or on-call operations at scale.
  • Experience with evaluation infrastructure for ML systems.
  • Experience driving infrastructure migrations in complex, multi-team environments.
  • Familiarity with monitoring and alerting tooling.

Skills

Technical proficiency in managing operational systems, incident response, and platform migrations. Strong communication and collaboration skills. Familiarity with AI safety principles and practices.

Benefits

Competitive compensation and benefits package, including optional equity donation matching, generous vacation and parental leave, flexible working hours, and a supportive office environment.

Pay

$290,000—$365,000 USD annually

Schedule

Hybrid schedule, with flexibility to work from home or in-office based on location and role requirements.

Similar jobs