Jobs · Engineering · California

Senior Staff Site Reliability Engineer

LiveRamp · San Francisco, CA · 2 wk ago
HybridEngineering$181k–$263k/yrFull-time

About the role

LiveRamp is looking for a Senior Staff Site Reliability Engineer who will set the technical direction for reliability engineering across LiveRamp's global infrastructure. This is a senior individual contributor role with organization-wide scope—you will define and own the SRE strategy, influence product and platform architecture decisions, and raise the engineering bar across multiple teams and regions.

Responsibilities

  • Define and own the SRE strategy across the organization—SLOs/SLAs, error budgets, and operational excellence frameworks
  • Oversee automation of critical areas to mitigate risk and align with engineering priorities
  • Develop and own some of the most complex software infrastructure spanning multiple products and services
  • Drive engineering-wide system design, automation, and performance optimization standards
  • Lead distributed systems architecture reviews and kickoffs across engineering teams
  • Drive high-quality API and interface designs across multiple teams
  • Drive overall architecture improvements across multiple products and services
  • Shape product and service design vision inside engineering, anticipating the unexpressed needs of internal teams
  • Understand global industry and market trends and apply them to deliver superior infrastructure solutions
  • Maintain a complete view of LiveRamp products and how SRE OKRs support the product roadmap
  • Contribute technical due diligence to M&A evaluations of potential acquisitions and partnerships
  • Serve as the escalation point of last resort for high-impact production incidents globally, leading postmortems with org-wide action items
  • Establish and enforce production readiness standards across engineering
  • Champion FinOps strategy across Kubernetes, cloud resources, and database infrastructure
  • Mentor Staff Engineers and provide technical feedback and guidance
  • Hold peers accountable for on-time, quality delivery
  • Represent LiveRamp's best interests in the broader technology ecosystem

Requirements

  • B.S./M.S. in Computer Science, Software Engineering, or equivalent
  • 10+ years in SRE, production engineering, or platform engineering; 3+ years at senior or staff level
  • Expert in Infrastructure as Code (Terraform) at scale across multi-environment, multi-team setups
  • Proven experience designing and operating highly available, globally distributed systems
  • Deep Kubernetes expertise: internals, autoscaling, multi-tenant workload management, and rightsizing
  • Advanced experience with real-time and NoSQL databases (SingleStore, ScyllaDB, Cassandra, DynamoDB)
  • Strong proficiency in Python and/or Go; able to build production-grade internal tooling adopted across teams
  • Expertise in observability engineering—SLOs, SLI pipelines, and high-signal alerting systems
  • Deep FinOps expertise: cost attribution, reserved capacity strategy, and cloud cost governance at scale
  • Experience maturing CI/CD platforms (Jenkins, CircleCI, or equivalent) for multiple engineering teams
  • Strong cloud security background: IAM, network segmentation, secrets management, SOC 2 / ISO 27001 (GCP and/or AWS)
  • Peer-recognized expert in a field relevant to LiveRamp's technology ecosystem
  • Exceptional communicator across both technical and executive audiences
  • Proven ability to lead without authority and influence across engineering organizations

Nice to have

  • Experience building or operating multi-region active-active architectures
  • Contributions to open source observability or infrastructure tooling
  • Experience with chaos engineering frameworks (Gremlin, Chaos Monkey, or equivalent)
  • Prior experience in a staff-plus IC role or as a technical lead for a global SRE organization
  • Familiarity with LLMs and AI-assisted development workflows, including tools such as Claude Code; experience applying agentic software development patterns to automate infrastructure tasks, incident response, or operational toil

Pay

The approximate annual base compensation range is $181,000 to $263,000. The actual offer, reflecting the total compensation package and benefits, will be determined by a number of factors including the applicant's experience, knowledge, skills, and abilities, geography, as well as internal equity among our team.

Benefits

  • Flexible paid time off, paid holidays, options for working from home, and paid parental leave
  • Comprehensive benefits package including medical, dental, vision, life and disability, an employee assistance program, and voluntary benefits
  • 401K matching plan—1:1 match up to 6% of salary
  • Employee Stock Purchase Plan - 15% discount off purchase price of LiveRamp stock (U.S. LiveRampers)
  • In-person and virtual events such as game nights, happy hours, camping trips, and sports leagues

Similar jobs