Jobs · Engineering · California

Senior Developer Technology Engineer - Edge Agentic AI

NVIDIA · Santa Clara, CA · 1 mo ago
EngineeringFull-time

About the role

As a Developer Technology Engineer, you will be at the forefront of innovation, working with leading industry partners and pioneering open-source projects to enable professional agentic AI workflows at the edge powered by NVIDIA RTX and DGX platforms.

Responsibilities

  • Work closely across internal engineering and product teams as well as external app developers and enterprise ISVs on solving local end-to-end agentic AI GPU deployment challenges on NVIDIA RTX & DGX.
  • Apply powerful profiling and debugging tools for analyzing most demanding accelerated end-to-end agentic AI workflows to detect insufficient system utilization resulting in suboptimal runtime performance.
  • Conduct hands-on trainings, develop sample code and host presentations to give good guidance on efficient end-to-end agentic AI deployment targeting optimal runtime performance.
  • Improve LLM & GenAI user experience by working on feature and performance enhancements of OSS software, including but not limited to projects like GGML, Llama.cpp, Ollama, vLLM, ONNX Runtime.
  • Collaborate with GPU driver and architecture teams as well as NVIDIA research to influence next generation GPU features by providing real-world workflows and giving feedback on partner and customer needs.
  • Provide technical leadership and mentorship to junior engineers, encouraging an inclusive and high-performing team environment.

Requirements

  • A proven track record of 5+ years of professional experience in local GPU deployment, profiling and optimization.
  • A Bachelor's or Master's degree or equivalent experience in Computer Science, Engineering, or a related field.
  • Strong proficiency in C/C++, Python, software design, programming techniques.
  • Familiarity with and development experience on Windows and Linux.
  • Experience with CUDA and NVIDIA's Nsight GPU profiling and debugging suite.
  • Some travel is required for conferences and for on-site visits with external partners.
  • Strong problem-solving skills and the ability to work both independently and collaboratively in a fast-paced environment.
  • Excellent interpersonal and communication skills and a passion for keeping track with the latest advancements in AI technology.

Qualifications

  • Experience with GPU-accelerated AI inference driven by NVIDIA APIs and SDKs, specifically TensorRT-RTX, cuDNN, NVIDIA Model Optimizer.
  • Expertise with professional agentic AI use cases, i.e., digital content creation and productivity workflows.
  • Experience working with open-source LLM and GenAI software.
  • Detailed knowledge of the latest generation GPU architectures.
  • Experience with AI deployment on NPUs and ARM architectures.

Benefits

We are widely considered to be one of the technology world’s most desirable employers, offering highly competitive salaries and a comprehensive benefits package. You will also be eligible for equity and benefits.

Pay

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is $152,000 - $241,500 for Level 3, and $184,000 - $287,500 for Level 4.

Schedule

Applications for this job will be accepted at least until July 26, 2026. This posting is for an existing vacancy.

Similar jobs