Principal Product Manager, Agentic Evals
About the Team
The Product Team creates high-quality end-to-end experiences for travelers, partners, and Expedia Group. Our focus on customer-centric innovation enables us to develop products that build loyalty and repeat business. We partner closely with teams across Expedia Group to drive growth and achieve results for our customers and the company.
Expedia Group powers travel for everyone, everywhere. As AI becomes core to how travelers plan, book, and manage their trips, our ability to evaluate quality, safety, and performance becomes essential. You will build the frameworks and tools that ensure every AI experience meets a high standard of clarity, accuracy, and trust.
The Core AI Experiences team defines the foundational methods for evaluating AI across the company. We create the metrics, tooling, and processes that guide how teams measure accuracy, safety, latency, helpfulness, and long-term traveler outcomes.
Responsibilities
- Define Expedia Group’s AI evaluation framework, principles, and adoption strategy
- Build tools and workflows that support reliable, scalable measurement across brands
- Partner with engineering and data science on metric design, instrumentation, and model evaluation
- Establish clear, durable standards for accuracy, safety, transparency, and performance
- Publish evaluation insights, scorecards, and recommendations that shape roadmaps
- Influence strategic decisions on AI investments and risk management
- Drive alignment across product, engineering, data science, and platform teams on evaluation strategy, standards, and prioritization
- Create governance models that ensure consistency and responsible deployment at scale
- Shape cross-team operating models and evaluation practices that guide how AI products are built and deployed across the company
Requirements
- Bachelor's degree in a technical, quantitative, or related field, or equivalent practical experience
- 10+ years of product management experience, including significant work with AI, ML, data platforms, or evaluation systems
- Strong technical depth and experience partnering with DS/ML teams
- Proven ability to build frameworks, metrics, or tools that scale across large organizations
- Experience with model evaluation, instrumentation, experimentation, or system reliability
- Strong written and verbal communication that brings clarity to complex technical spaces
- Track record of influencing cross-functional leaders and driving alignment
Preferred Qualifications
- Experience building evaluation platforms, ML tooling, or governance frameworks
- Familiarity with LLM behavior, model evaluation methods, or prompt testing
- Understanding of trustworthy AI principles, safety evaluation, and failure-mode analysis
Pay
The total cash range for this position in Seattle is $224,000.00 to $313,500.00. Employees in this role have the potential to increase their pay up to $358,500.00, which is the top of the range, based on ongoing, demonstrated, and sustained performance in the role.
The total cash range for this position in San Jose is $242,000.00 to $338,500.00. Employees in this role have the potential to increase their pay up to $387,000.00, which is the top of the range, based on ongoing, demonstrated, and sustained performance in the role.
Starting pay for this role will vary based on multiple factors, including location, available budget, and an individual’s knowledge, skills, and experience.
Benefits
- Medical, dental, and vision coverage
- Paid time off
- Employee Assistance Program
- Wellness and travel reimbursement
- Travel discounts
- International Airlines Travel Agent Network (IATAN) membership
Learn more about life at Expedia Group at https://careers.expedigroup.com/life.