Senior Solutions Architect, NVIDIA Cloud Partners
NVIDIA · Texas, United States · 4 wk ago
RemoteRemoteSalesFull-time
About the role
NVIDIA is seeking an experienced Solutions Architect to bridge design to deployment of large-scale AI/HPC GPU infrastructure. The role involves driving adoption and consumption by integrating libraries, frameworks, models, and software applications. It also includes delivering GenAI, AI, and ML hardware/software to production with key customers and partners, supporting partners in building next-gen GPU platforms.
Responsibilities
- Collaborate with NVIDIA Cloud Partners to create, implement, and deliver innovative hardware and software solutions.
- Partner with Solution Architects, Account Managers, Engineering, Product, and business leaders to align on strategies, assess technical needs, and secure business opportunities for NVIDIA.
- Become the primary technical driver for customers during the design, development, construction, integration, and production of GPU Cloud infrastructure and applications throughout the entire customer lifecycle.
- Conduct regular technical customer meetings for project/product details, feature discussions, introduction to new technologies, and debugging sessions.
- Work closely with customers to build and adopt NVIDIA solutions, including Proof of Concepts (PoCs) to address critical business needs covering infrastructure, libraries, and applications.
- Prepare and deliver technical content to customers, including presentations, workshops, reference architectures, tutorials, and publications.
Requirements
- BS/MS/PhD in Electrical/Computer Engineering, Computer Science, Physics, Mathematics, or other Engineering fields or equivalent experience.
- 10+ years of Solution Engineering (or similar Sales Engineering, Cloud Engineering, Solution Architecture) including experience working directly with partners and customers.
- Experience crafting and deploying large-scale cluster environments, hands-on experience designing, developing, delivering distributed Cloud architectures.
- Strong fundamentals in programming, optimizations, and software design, especially in Python and Deep Learning frameworks such as PyTorch and TensorFlow.
- Practical expertise in fine-tuning and deploying models, integrating software application stacks, libraries, and frameworks to drive consumption from GPU platforms.
- Motivation and skills to own and drive complex multi-disciplinary technical engagements with customers throughout the full customer lifecycle and cross-functional teams.
- Efficient time management and capable of balancing multiple tasks.
- Excellent presentation, communication, and collaboration skills.
- Self-starter with a passion for growth, continuous learning, and sharing insights.
Qualifications
- Practical experience with NVIDIA GPUs, software libraries, frameworks, and foundation models, such as NVIDIA Nemotron, NVIDIA NeMo Framework, NVIDIA Dynamo, NeMo Retriever, NVIDIA Triton Inference Server, TensorRT, TensorRT-LLM, NVIDIA CUDA-X.
- Hands-on expertise with scaled AI cloud environments (e.g., AWS, Azure, GCP) and on-premises/hybrid infrastructure, in particular inference and training workloads.
- Familiarity with NVIDIA hardware (such as GPUs, networking, storage) and systems technology such as NCCL, DCGM, UFM, Mission Control, Base Command Manager.
- Proficiency with large-scale AI model training/deployment encompassing GPU systems, performance testing, AI benchmarking, fine tuning, strong focus on MLOps and cluster orchestration (SLURM, K8s, orchestrator, load balancing, cloud architecture).
- Experience working with enterprise developers and strong customer-facing skills.
Benefits
Base salary will be determined based on location, experience, and the pay of employees in similar positions. The base salary range is $184,000 - $287,500. Eligible for equity and benefits.
Pay
$184,000 - $287,500
Schedule
Full-time