Software Development Engineer, Neuron Foundation Tools
About the role
As the Software Development Engineer for the Neuron Foundation Tools Team, you will develop and maintain high-performance monitoring and profiling tools for machine learning applications and AI accelerators. You will work on the design, development, and deployment of the Neuron Profiler and other Neuron Tools, which help internal and external customers optimize AI workloads across hardware platforms such as Trainium and Inferentia devices by providing deep insights into performance bottlenecks and system behavior.
In this role, you will manage the full development life cycle of the Neuron Profiler/Tools toolchain, ensuring scalability, reliability, and usability. You will collaborate with cross-functional teams to ensure the C++ compiler and runtime generate key information for customers to understand and optimize performance on custom hardware. You will also drive innovations to support multiple frameworks, such as PyTorch, JAX, and XLA.
Responsibilities
- Develop and maintain high-performance monitoring and profiling tools for machine learning applications and AI accelerators.
- Design, develop, and deploy the Neuron Profiler and other Neuron Tools.
- Manage the full development life cycle of the Neuron Profiler/Tools toolchain, ensuring scalability, reliability, and usability.
- Collaborate with cross-functional teams to ensure the C++ compiler and runtime generate key performance insights for customers.
- Drive innovations to support multiple frameworks (e.g., PyTorch, JAX, and XLA).
- Work with executive leadership and senior management to define and deliver product directions.
- Improve the performance of ML kernels and frameworks.
Requirements
- 3+ years of non-internship professional software development experience.
- 2+ years of non-internship design or architecture experience (design patterns, reliability, and scaling) of new and existing systems.
- Experience programming with at least one software programming language.
Qualifications
- 3+ years of full software development life cycle experience, including coding standards, code reviews, source control management, build processes, testing, and operations.
- Established background in building AI/ML and performance analysis tools.
- Experience with ML-specific profiler tools (e.g., PyTorch Profiler or TensorFlow Profiler).
- Direct customer-facing experience and a strong motivation to achieve results.
- Bachelor's degree in computer science or equivalent.
Benefits
- Comprehensive health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance, and optional supplemental life plans).
- Employee Assistance Program (EAP) and Mental Health Support.
- Medical Advice Line and Flexible Spending Accounts.
- Adoption and Surrogacy Reimbursement coverage.
- 401(k) matching.
- Paid time off and parental leave.
- Sign-on payments and restricted stock units (RSUs).
Pay
Base salary range: $143,700.00 - $194,400.00 USD annually (USA, WA, Seattle). Final compensation will be determined based on experience, qualifications, and location.
About the team
Our team values work-life balance, focusing on the flow between personal and professional life to bring energy to both. We offer flexibility in working hours to help you find your own balance.
We are dedicated to supporting new members with a broad mix of experience levels and tenures. Our environment celebrates knowledge sharing and mentorship, and we prioritize your career growth by assigning projects that help you develop into a well-rounded professional.
We embrace diversity and inclusion, with ten employee-led affinity groups reaching 40,000 employees globally. Our culture is reinforced by Amazon’s 16 Leadership Principles, encouraging diverse perspectives, curiosity, and trust.