Jobs · Engineering · California

Senior DevOps Engineer - E-commerce

NVIDIA · Santa Clara, CA · 2 wk ago
HybridEngineeringFull-time

About the role

We are looking for an outstanding DevOps and Site Reliability Engineer to join the NVIDIA e-commerce team. You will be a key architect of our e-commerce platform, ensuring that our systems are scalable, resilient, and automated.

Responsibilities

  • Architect and refine automated deployment Jenkins pipelines to ensure seamless, zero-downtime releases.
  • Design, build, and maintain enterprise-scale infrastructure using Terraform.
  • Establish modular, reusable patterns for AWS resources.
  • Optimize and manage sophisticated AWS environments with a focus on cost-efficiency and security.
  • Transition our monitoring from reactive to proactive using AI-powered observability tools (e.g., Datadog Watchdog) for automated root cause analysis (RCA) and anomaly detection.
  • Define and monitor Service Level Objectives (SLOs) and Service Level Agreements (SLAs).
  • Lead incident response and conduct thorough post-mortems to improve system resilience.

Requirements

  • 8+ years or equivalent industry experience
  • Bachelor's/Master's Degree in Computer Science, Software Engineering, or equivalent experience
  • Exceptionally strong background in developing CI/CD processes and deployment pipelines using Jenkins
  • Extensive experience architecting on AWS Cloud and running services such as API Gateway, Lambda, EKS/ECS, RDS, S3, and SQS
  • Expert-level knowledge of Terraform (including state management, workspaces, and complex module development)
  • Advanced experience with Kubernetes (EKS) and Docker, including orchestration, service meshes, and Helm
  • Strong proficiency in a scripting language, such as Python, for automation and custom tooling
  • Strong communication skills

Qualifications

  • Deep understanding of DNS and CDNs (e.g., Akamai, CloudFront)
  • Demonstrated use of AI tools to improve productivity and the quality of releases
  • Applies secure-by-design principles across infrastructure, deployment automation, and operational processes

Skills

  • Terraform Expert
  • Expert-level knowledge of Kubernetes (EKS)
  • Advanced experience with Docker
  • Strong proficiency in a scripting language, such as Python
  • Strong communication skills

Benefits

  • Your base salary will be determined based on your location, experience, and the pay of employees in similar positions.
  • The base salary range is $176,000 - $276,000 for Level 4, and $208,000 - $333,500 for Level 5.
  • You will also be eligible for equity and benefits.

Pay

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is $176,000 - $276,000 for Level 4, and $208,000 - $333,500 for Level 5.

Schedule

Applications for this job will be accepted at least until July 18, 2026.

Ways to Stand Out

  • Deep understanding of DNS and CDNs (e.g., Akamai, CloudFront)
  • Demonstrated use of AI tools to improve productivity and the quality of releases
  • Applies secure-by-design principles across infrastructure, deployment automation, and operational processes

Company Information

  • NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer.
  • We highly value diversity in our current and future employees.
  • We do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Similar jobs