Jobs · Quality Assurance

AI Enablement & Governance- AI Quality & Evaluation Lead

Ladders · United States · 4 wk ago
RemoteRemoteQuality Assurance$150k–$200k/yrFull-time

Responsibilities

  • Partner with AI Engineers and Data Scientists during design to set quality acceptance criteria
  • Embed quality standards into model architecture from the beginning, particularly for complex AI systems
  • Define golden truth and dataset standards, ensuring they reflect real-world complexities
  • Establish quality expectations for third-party AI models to maintain governance standards
  • Develop a structured evaluation framework to assess AI systems against quality metrics like Goodness-of-Fit
  • Create automated performance metrics for Generative AI, including detection of hallucinations and completeness
  • Set up ongoing monitoring for model performance including triggers for necessary interventions

Qualifications

  • 5-8 years of experience in Data Science, ML Engineering, or AI Quality with a focus on evaluation and statistical validation
  • Practical experience in designing RAG, LLM-based Agents, or traditional ML pipelines alongside engineers
  • Expert-level Python skills, particularly with tools like Pandas and Scikit-learn, and familiarity with evaluation frameworks like RAGAS or MLflow
  • Experience in influencing stakeholders to adopt and implement quality practices across various teams
  • Strong communication skills that connect governance policies with technical implementation nuances
  • Bachelor's degree in a technical field or equivalent experience

Benefits

  • Opportunity to impact the AI lifecycle through quality and evaluation leadership
  • Collaborative environment working alongside AI Engineering and Data Science teams
  • Exposure to advanced AI technologies, including RAG systems and Generative AI
  • Gaining experience in governance frameworks for AI that are increasingly important in the industry

Similar jobs