Distinguished Engineer, End-to-End Scaling Performance Architecture
About the role
Join NVIDIA's architecture organization to define how future accelerated computing systems scale from a single processor to multi-die, multi-GPU, and multi-node platforms. In this role, you will set long-term performance strategy across applications, systems, and architecture, including DRAM, NVLink, and chip-to-chip (C2C) interconnects. The role focuses on architectural direction and application outcomes. You will identify where data movement, communication, memory behavior, topology, and compute limit scaling, then turn those insights into priorities that guide multiple product generations. Domain teams own detailed implementation and delivery; you will align their decisions around a shared end-to-end strategy so local improvements create meaningful system-level gains.
What you will be doing
- Define the multi-generation strategy for application scaling across DRAM, NVLink, C2C, compute, and the supporting software stack
- Translate the behavior of important AI, HPC, and accelerated computing applications into architectural requirements, performance targets, and investment priorities
- Build a clear view of how bottlenecks shift as workloads scale across dies, GPUs, nodes, model sizes, data sets, and communication patterns
- Evaluate system-level trade-offs across bandwidth, latency, capacity, topology, coherence, power, area, cost, programmability, and resiliency
- Establish common workload scenarios, scaling metrics, models, and decision frameworks so architecture teams can compare proposals against application outcomes
- Identify architectural discontinuities and emerging technology opportunities early enough to shape product and technology decisions
- Partner with DRAM, NVLink, C2C, GPU, CPU, system, and software architects to align around shared performance limits and high-value opportunities
- Work with application, framework, compiler, runtime, modeling, and post-silicon teams to connect measured behavior with future architecture choices
- Provide clear recommendations to senior technical and business leaders, including assumptions, sensitivities, risks, and expected impact
- Mentor system performance architects, strengthen technical communities across teams, generate sustained intellectual property, and help influence the direction of large-scale accelerated computing
What we need to see
- MSEE, MSCE, PhD, or equivalent experience in Electrical Engineering, Computer Engineering, Computer Science, or a related field
- 18+ years of relevant industry or academic experience, including experience setting architecture direction for complex, high-performance systems
- Deep understanding of system performance and scaling, including interactions among DRAM behavior, high-bandwidth fabrics such as NVLink, and C2C communication
- Strong application-level intuition, including the ability to connect workload algorithms, parallelism, communication, locality, and data movement to architecture choices and measurable outcomes
- Experience with workload characterization, analytical or simulation-based performance modeling, bottleneck analysis, and architecture trade-off evaluation
- Record of identifying cross-domain opportunities that may not be visible when teams optimize individual components separately
- Demonstrated ability to create and advance a multi-generation technical strategy through influence across silicon, systems, software, and application teams
- Clear communication and sound judgment in ambiguous technical areas, with the ability to explain complex system trade-offs to specialists and executive leaders
- Experience mentoring senior engineers into broader architecture leadership roles and building strong technical communities
Ways to stand out from the crowd
- Shaped product or technology roadmaps around application-level scaling needs
- Helped architecture teams build a quantitative understanding of bottlenecks across memory, interconnect, compute, and software
- Identified high-impact trade-offs early enough to guide product, architecture, or technology investment
- Delivered measurable end-to-end improvements in performance, efficiency, or scaling for priority applications
Pay
Base salary range: 320,000 USD - 488,750 USD, determined based on location, experience, and pay of employees in similar positions. Equity and benefits also provided.