Jobs · Engineering · California

Senior Cloud Software Engineer, DGXC Data Services

NVIDIA · Santa Clara, CA · 1 mo ago
EngineeringFull-time

About the role

The NVIDIA DGXC Data Services team builds cloud-native systems, frameworks, and services for managing data across hybrid and multi-cloud infrastructure. We are building the next-generation data and storage infrastructure to solve some of the hardest problems in AI: storage, access, ingestion, governance, observability, and data management for exabyte-scale, high-performance GPU-based training and inference jobs. Our work gives NVIDIA teams the foundational capabilities they need to build, train, deploy, and operate AI products at scale without reinventing critical data infrastructure for every workload.

Responsibilities

  • Build cloud-native data and storage services for hybrid and multi-cloud infrastructure, including dataset discovery, ingestion, governance, checkpointing, observability, and low-latency access.
  • Develop scalable cloud-native services and APIs that support exabyte-scale, high-performance GPU training and inference workflows.
  • Work closely with product managers, internal AI teams, platform teams, and partner engineering teams to understand requirements and turn them into reliable production systems.
  • Collaborate with SRE, operations, and support teams to improve service reliability, performance, observability, on-call readiness, and operational scale.
  • Use modern software engineering practices, including AI-assisted and agentic development workflows, while maintaining high standards for design, testing, security, and verification.

Requirements

  • BS in Computer Science, Information Systems, Computer Engineering, or equivalent experience, with 5+ years of software engineering experience.
  • Strong foundation in algorithms, data structures, distributed systems, and practical software design.
  • Experience building, shipping, and operating backend or cloud-native services using Kubernetes, cloud providers such as AWS, GCP, or Azure, and languages such as Go, Python, Rust, C/C++, or Java.
  • Ability to design APIs, document systems, reason through tradeoffs, communicate clearly, and break ambiguous problems into practical execution plans.
  • Experience working across engineering, product, platform, and operations teams to deliver reliable production software.
  • Curiosity and practical judgment around AI-assisted or agentic engineering workflows, including using clear intent, specifications, acceptance criteria, tests, and verification to guide development.

Qualifications

  • Hands-on experience building, scaling, or operating large-scale data, storage, or ML infrastructure services.
  • Experience solving enterprise-grade data management, governance, analytics, or AI workflow problems with modern data and ML infrastructure technologies.
  • Strong background in distributed systems, storage systems, cloud infrastructure, performance engineering, observability, or agentic engineering practices.

Skills

  • Strong foundation in algorithms, data structures, distributed systems, and practical software design.
  • Experience building, shipping, and operating backend or cloud-native services using Kubernetes, cloud providers such as AWS, GCP, or Azure, and languages such as Go, Python, Rust, C/C++, or Java.
  • Ability to design APIs, document systems, reason through tradeoffs, communicate clearly, and break ambiguous problems into practical execution plans.
  • Experience working across engineering, product, platform, and operations teams to deliver reliable production software.
  • Curiosity and practical judgment around AI-assisted or agentic engineering workflows, including using clear intent, specifications, acceptance criteria, tests, and verification to guide development.

Benefits

NVIDIA offers competitive compensation packages, including base salaries ranging from $152,000 to $241,500 for Level 3 and $184,000 to $287,500 for Level 4, depending on experience and location. Additionally, employees are eligible for equity and comprehensive benefits.

Pay

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is $152,000 - $241,500 for Level 3, and $184,000 - $287,500 for Level 4.

Schedule

We are currently accepting applications for this position until July 10, 2026.

Ways to Stand Out

  • Hands-on experience building, scaling, or operating large-scale data, storage, or ML infrastructure services.
  • Experience solving enterprise-grade data management, governance, analytics, or AI workflow problems with modern data and ML infrastructure technologies.
  • Strong background in distributed systems, storage systems, cloud infrastructure, performance engineering, observability, or agentic engineering practices.

Equal Opportunity Employer

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Similar jobs