Senior ML Platform Engineer - AD/ADAS
Woven by Toyota · Palo Alto, CA · 1 mo ago
HybridInformation Technology$140k–$230k/yrFull-time
Responsibilities
- Design, build, maintain, optimize and support the ML Platform’s systems and tools for perception, prediction, and planner development.
- Allowing numerous ML engineers to effectively & efficiently iterate on dataset curation, ML modeling, training, evaluation and deployment of ML models into our functionally safe AD/ADAS stack, shipped in millions of Toyota vehicles.
- Develop user-friendly tooling, frameworks and libraries to support the overall ML engineering effort, from ML modeling, to tracking performance metrics and introspecting failure modes.
- Build and maintain efficient dataset generation, cloud training and evaluation pipelines.
- Develop and review code with other ML and ML Platform engineers to facilitate rapid incremental improvements.
- Optimize the current processes, tooling and supporting infrastructure to accelerate the overall ML engineering effort, and contribute to the long term strategy for several of our systems and products.
- Work in a high-velocity environment and employ agile development practices.
- Work in a hybrid workspace, with the requirement to be present in our Palo Alto (USA) office three days per week.
Requirements
- BSc / BEng (MS / PhD nice-to-have) in Machine Learning, Computer Science, Robotics or related quantitative fields, or equivalent industry experience.
- 5+ years of experience with data structures, algorithms, design patterns, and software engineering best practices.
- 2+ years of experience with UNIX-based systems (Linux or similar), Python, and PyTorch/Tensorflow.
- 2+ years of experience in the full MLOps cycle covering data cleansing, data sampling, data curation, pre-processing, efficient data loading, distributed training, testing, evaluation, deployment, inference optimization and deployment in the cloud and on edge compute platforms.
- Experience with Docker and CI systems such as GitHub Actions.
Qualifications
- Business-level proficiency in English, able to write technical documents (e.g., for software documentation).
- Nice to haves include:
- 2+ years of experience with Apache Spark, Airflow, Flyte, Flink, Ray, or similar ML pipelines technologies.
- 2+ years using modern systems programming languages (e.g., Rust and/or C++) and a modern build system (preferably Bazel), and systems-level debugging knowledge, in a professional environment.
- Experience with SIMD/SIMT parallelism, GPU programming, multithreading.
- Experience with Terraform, AWS, Observability, and Kubernetes in production.
- Experience with Google Big Query, Snowflake or AWS Redshift in production.
- Experience in optimizing deep-learning models towards specific hardware targets.
- Experience in self-driving, robotics, computer vision, or motion planning.
- Experience working in a fast-paced environment, collaborating across teams and disciplines.
- Business-level proficiency in Japanese.
Benefits
- Excellent health, wellness, dental and vision coverage
- A rewarding 401k program
- Flexible vacation policy
- Family planning and care benefits
Pay
The base pay for this position ranges from $ 140,000 - $ 230,000 a year. Your base salary is one part of your total compensation. We offer a base salary, short term and long term incentives, and a comprehensive benefits package. The total compensation offered to an employee will be dependent upon the individual's skills, experience, qualifications, location, and level.