Jobs · Engineering · California

Site Reliability Engineer

Obsidian Security · Palo Alto, CA · 1 wk ago
On-siteEngineering$165k/yrFull-time

Obsidian Security is the leading SaaS security platform, trusted by global enterprises like Snowflake, T-Mobile, and Algolia. We protect 200+ organizations across North America, Europe, the Middle East, Southeast Asia, Australia, and New Zealand, including many of the world’s largest Fortune 1000 and Global 2000 companies. Founded in 2017 and backed by top investors like Greylock, Obsidian was built to close a critical gap: securing SaaS apps where business happens—Microsoft 365, Salesforce, and hundreds more. The company offers a complete SaaS security platform to reduce risk, detect and respond to threats, and prevent breaches at the source. Obsidian was built by leaders who redefined endpoint and identity security at CrowdStrike, Okta, Cylance, and Carbon Black. Now, they’re transforming how SaaS is secured. With AI driving rapid SaaS growth and complexity, agentic AI tools gain privileged access to sensitive data through integrations, creating new risks most security tools miss. Obsidian uniquely detects anomalous OAuth token activity and manages integration risks.

About The DevOps / SRE Team

The DevOps/SRE team at Obsidian ensures that engineering excellence translates into stable, scalable, and high-performing production systems. We work closely with Engineering, Quality Engineering, and Customer Support to deliver end-to-end services that bring code to life and maintain our world-class SaaS security platform.

Responsibilities

  • Support and maintain the service quality of our customer-facing SaaS security platform
  • Address complex challenges around scalability, reliability, observability, and cost efficiency
  • Collaborate with Engineering teams to maintain and enhance Helm charts, application deployment, monitoring and CI/CD pipelines
  • Embed into the engineering team so that you understand the application deeply
  • Define service verification strategies and implement them as part of the CI/CD process to meet SLAs
  • Improve developer experience by optimizing CI/CD workflows and performance
  • Participate in the on-call rotation, providing 24/7 support in coordination with our global SRE team
  • Monitor, debug, and optimize production infrastructure and services on AWS/GCP

Requirements

  • 3+ years of experience in a DevOps or SRE role supporting SaaS services on GCP and/or AWS
  • Bachelor's degree in Computer Science or related field
  • Strong proficiency in Kubernetes, microservices architecture, Helm, GitLab CI/CD, and ArgoCD, Prometheus, Grafana
  • Programming experience in at least one language; Golang or Python preferred
  • Deep understanding of autoscaling, version upgrades, and cloud service optimization
  • Bonus if you're familiar with technologies like Kafka, Elasticsearch, PostgreSQL, ScyllaDB, Databricks, Dagster, Sentry, Kong

Benefits

  • Competitive compensation with equity and 401k
  • Comprehensive healthcare with dental and vision coverage
  • Flexible paid time off and paid holiday time off
  • 12 weeks of new parent or family leave
  • Personal and professional development resources

Pay

Base Salary Range: $165,000 USD - $190,000 USD. The base pay will vary based on factors such as work location, as well as the knowledge, skills, and experience of the candidate. In addition to a competitive base salary, this position is eligible for equity awards and may be eligible for sales commission or incentive compensation based on the role or function within the company.

Similar jobs

Site Reliability Engineer

Cracker BarrelTennessee, United States· 1 mo ago
Engineeringapply on cbrlgroup.wd503.myworkdayjobs.com