Jobs · Engineering · Washington

Senior Software Engineer, GoLang - DSX MaxQ

NVIDIA · Redmond, WA · 3 wk ago
EngineeringFull-time

About the role

NVIDIA is seeking talented software engineers to contribute to the development of enterprise GPU management and monitoring tools. This role involves designing and building cloud-native management agents, Kubernetes integrations, and end-to-end integration solutions that integrate GPUs with the rest of the datacenter software management ecosystem.

Responsibilities

  • Develop and maintain distributed, robust, and scalable Go programs deployed to Kubernetes environments that manage large datacenters
  • Develop and maintain user-space applications, containers, Go-bindings, and CLI tools
  • Enable GPU management integration with the state-of-the-art open-source ecosystem, including Kubernetes and Docker
  • Support internal and external users through bug fixes, documentation, and feature improvements
  • Maintain high-quality products through robust test coverage

Requirements

  • BS or higher in Computer Science or equivalent experience
  • 5+ years of meaningful industry experience with a strong Go and Kubernetes development background
  • User space development and debugging expertise in Linux environments
  • Experience with APIs and interface design
  • Outstanding written and verbal interpersonal skills
  • Business level English
  • Strong motivation and commitment to learn new skills
  • Ability to execute all aspects of the software development lifecycle
  • Ability to manage time in a fast, heavily multitasked environment
  • Development experience with Rust, Python and/or C, C++
  • Experience with distributed systems and concurrent applications, especially in a Kubernetes environment
  • Experience developing and maintaining enterprise software
  • Experience deploying, managing, and debugging applications in a Kubernetes environment

Qualifications

  • Background with containers (e.g. Docker, OCI), orchestration frameworks, and logging/telemetry backends with Kubernetes
  • Monitoring stacks with tools such as Prometheus, Loki and Grafana
  • Experience with modern UI development in React and Node.js or similar frameworks
  • Experience developing Kubernetes operators or Helm charts
  • Experience with HPC job schedulers like Slurm or Run.AI
  • Familiarity with Kubernetes internals
  • Experience with GPU programming with CUDA
  • Experience with Jenkins and GitHub/GitLab CI/CD pipelines

Skills

  • Linux background
  • Familiarity with modern cloud-native systems
  • Proven work ethic

Benefits

  • Base salary range: $152,000 - $241,500 for Level 3, and $184,000 - $287,500 for Level 4
  • Eligibility for equity and benefits

Pay

  • Base salary determined based on location, experience, and pay of employees in similar positions

Schedule

  • Dynamic work environment with many exciting opportunities

Similar jobs