Jobs · Engineering · Colorado

Sr. Machine Learning Engineer

AppFolio · Denver, CO · 1 wk ago
EngineeringFull-time

Who We Are Looking For

We're hiring a Senior Machine Learning Engineer to design and ship the next generation of voice and conversational AI agents within Realm-X. This role helps define AppFolio's production voice and chat agent pipelines, working at the intersection of LLM agent frameworks, real-time voice technology, and streaming infrastructure.

Your Impact

  • Ship Voice & Text Agents: Architect and ship voice and text agent pipelines that handle real-time, multi-turn customer interactions.
  • Reasoning vs. Latency: Make principled trade-offs between reasoning depth and latency across frontier LLMs, smaller models, and routing strategies.
  • Lead a Pod: Lead a small pod of ML and platform engineers; raise the bar on agent evaluation, observability, and incident response.
  • Define Quality: Partner with Product and Voice channel teams to define KPIs, eval harnesses, and acceptance criteria for agent quality.
  • Optimize for Voice: Drive selective Small Language Model (SLM) fine-tuning and inference optimization for voice latency and cost.

Qualifications

  • You have shipped production AI agents serving real users in voice and/or text channels.
  • You think in pipelines and systems, not just models.
  • You move fast, deliver impact, and maintain sound engineering judgment.
  • You are humble, collaborative, and low-ego, and you elevate those around you.
  • You value work-life balance as a foundation for sustained high performance.

Must Have

  • Agent frameworks: Deep, shipped experience with LangChain, LangGraph, LangSmith, and LangChain Deep Agents (or equivalent agent frameworks).
  • Voice stack: Hands-on with Voice-to-Voice models and traditional TTS / STT pipelines; understands the trade-offs between end-to-end voice models and modular STT → LLM → TTS architectures.
  • LLM fluency: Strong grasp of LLM reasoning behavior, tool use, structured output, and reasoning-vs-latency trade-offs across providers.
  • Telephony & cloud: Production experience with Twilio (or comparable telephony) and AWS.
  • Engineering: Expert Python, async programming, and WebSockets for real-time, bidirectional streaming.
  • ML fundamentals: Solid foundation in deep learning, model evaluation, and inference optimization; able to deploy with Docker on AWS.
  • Leadership: Demonstrated ability to lead a small team, mentor engineers, and partner credibly with Product and Design.

Nice to Have

  • Fine-tuning Small Language Models for domain-specific voice applications.
  • Familiarity with RAG over structured business data and tool-using agents over API surfaces.
  • Prior experience in regulated or customer-facing industries with strict reliability requirements.
  • Publicly verifiable work on GitHub, in open-source agent frameworks, or in community competitions.

Similar jobs