Senior AI/ML Engineer
About the role
Infinity Loop is a venture-backed AI SaaS startup focused on helping companies manage their third-party vendor relationships more effectively. The company's platform uses advanced AI models to generate negotiation strategies and provide legal risk insights, aiming to save customers money and reduce risk during renegotiations.
Responsibilities
- Architect the AI stack – selecting models, vector stores, and cloud infrastructure to lay the groundwork for a scalable and secure AI pipeline
- Own the entire pipeline – from data ingestion and cleaning to model deployment and monitoring
- Operationalize AI – implementing CI/CD for models and data, reproducible experiments, drift alerts, and usage dashboards
- Optimize for enterprise realities – addressing issues such as latency, cost, PII handling, and compliance-friendly logging
- Collaborate closely with the Head of Engineering and product stakeholders to prioritize and execute high-impact work
- Make pragmatic decisions in ambiguous environments, balancing technical elegance with business needs
- Help shape the engineering culture, tooling, and processes at an early-stage company
Requirements
We are seeking an experienced, hands-on engineer with a strong foundation in machine learning, infrastructure, and product development. This role requires a combination of technical expertise and a builder’s mindset, with a focus on shipping fast, troubleshooting deeply, and scaling successful solutions.
- 7+ years of professional experience, including 5+ years of experience writing clean, tested, production-ready code, particularly in building distributed, cloud-native services
- Hands-on experience building and deploying sophisticated LLM pipelines in real-world applications, including experience with tokenization, context windows, sampling/temperature settings, and failure modes like hallucinations or drift
- Experience with generative AI or NLP products that balance real-world constraints like latency, cost, and data privacy
- Expertise in LLM tuning and retrieval, including prompt crafting, RAG (retrieval-augmented generation), embeddings, fine-tuning, and hallucination detection/mitigation
- Broad ML/data science experience, turning raw data into models (e.g., classification, ranking, or NLP use cases)
- ML Ops mindset, with experience in CI/CD for models, experiment tracking, and production monitoring
- Product instincts - working closely with domain experts, prototyping quickly, and improving through iteration
- Technical leadership skills, including setting engineering standards, writing clear documentation, and mentoring others
- A bias for action - unblocking oneself, shipping what matters, and polishing to completion
- Proficiency in ML/LLM techniques, tools, and libraries, able to write and explain non-trivial code during interviews and in daily work
- Excellent communicator in design docs, diagrams, and PRs, with the ability to explain trade-offs across engineering and product teams
Qualifications
Strong written and verbal communication skills are essential, as you will need to clearly explain design decisions, trade-offs, and system behavior to both technical and non-technical stakeholders.
Skills
The ideal candidate should have a strong background in machine learning, infrastructure, and product development, with a focus on practicality and innovation.
Benefits
Infinity Loop offers competitive compensation and early-stage equity to its employees, providing significant ownership and growth potential.
Pay
Details on pay will be provided during the interview process.
Schedule
Full-time position available.