Jobs · Engineering · California

Senior AI Researcher

Ivo · San Francisco, CA · 2 days ago
On-siteEngineering$200k–$325k/yrContract

Why Join Ivo?

The way agreements are reviewed hasn't changed in four thousand years - a human reads the whole thing and tries not to miss anything. Ivo is the contract intelligence platform of choice for companies like Uber, Meta, Canva, IBM, and Shopify. We recently raised our Series B and have grown 800% over the last 12 months.

Engineering at Ivo

  • Engineers At Ivo Are Inventors. Ivo Was First-to-market With An AI agent that lives in MS Word and edits the document for you [2023]
  • Ditching imprecise embeddings models in favor of agentic RAG [2023]
  • Large-scale LLM-based legal fact extraction [2024]
  • A legal assistant that can search large contract databases without sacrificing accuracy [2024]
  • Clustering legal documents descended from the same family [2025]
  • Automatic deviation analysis to locate buried risk in huge contract databases [2025]
  • Merging contracts with their amendments to produce a time series of "composite" contracts (a customer actually cried when we showed her this) [2025]

The Role: Why, What, and Who

AI Researchers are the engine of innovation at Ivo. You'll push the state of the art in applying LLMs, deep learning, and advanced AI techniques to the unique and high-stakes challenges of the legal domain — where accuracy, explainability, and robustness aren't nice-to-haves, they're the product. Your research will translate directly into features that redefine how legal professionals work. This isn't an academic role. It's a chance to see your research fundamentally change an industry.

What You'll Do

  • Own a research roadmap end-to-end: identifying the right problems, designing experiments, prototyping, and shipping the winners into production alongside the engineering team.
  • Advance the core AI platform. Design and implement novel approaches to the problems at the heart of Ivo's product: reasoning over long-context legal corpora, contract comparison and redlining, information extraction, and automated drafting and editing.
  • Make our models trustworthy. Conduct research on procedural hallucination detection and resolution, calibration, and explainability.
  • Push frontier techniques into production. Explore and apply advanced fine-tuning, PEFT, and distillation techniques to make our models faster, cheaper, and more accurate on legal-specific tasks.
  • Evaluate emerging work in agentic systems, long-context modeling, and reasoning, and figure out which ideas actually move the needle for our customers.
  • Build the evaluation infrastructure. Design and maintain datasets, benchmarks, and evals for training and measuring model performance on complex legal text. Define the metrics that matter, and hold the team to them.
  • Ship. Partner closely with Engineering and Product to take prototypes from notebook to production, write internal reports that influence the technical direction of the platform, and present findings to both technical and non-technical audiences across the company.

Who you are

  • Required:
  • A Ph.D. in Computer Science, Engineering, Mathematics, Physics, or a related quantitative field — or equivalent industry research experience with a comparable track record.
  • Evidence of exceptional ability — a paper, shipped system, open-source contribution, competition result, or hard problem you cracked that puts you meaningfully ahead of your peers.
  • Deep, hands-on experience in deep learning research and development, particularly with LLMs.
  • Strong working knowledge of modern frameworks (PyTorch, JAX, or TensorFlow) and the surrounding open-source ecosystem.
  • Expertise in at least one of: agentic systems, reasoning, parameter-efficient fine-tuning (PEFT) methods, quantization, inference optimization (e.g., speculative decoding), hallucination mitigation, novel architectures in deep learning, or robust evaluation methodology for LLMs.
  • Excellent communication skills, with the ability to articulate complex research findings clearly to both technical and non-technical audiences.
  • A bias toward action: you ship rather than perfect, you measure rather than guess, and you'd rather have a working prototype today than a polished plan next week.
  • Nice To Have:
  • Publications at top venues — e.g., ML/AI conferences (NeurIPS, AAAI, ICML, ICLR), or peer-reviewed journals in adjacent quantitative fields (Journal of Computational Physics, SIAM journals, Journal of the ACM, Nature, Science, PNAS) — or equivalent strong research contributions in industry.
  • Experience with long-context modeling, retrieval, or grounded generation in high-stakes domains (legal, medical, financial).
  • Prior work on hallucination detection, calibration, or interpretability.
  • A track record of building from zero in fast-paced startup or research environments.

What We Offer

  • Competitive Compensation: Final offer details are determined based on experience, expertise, and overall fit.
  • Equity: Meaningful ownership in a company that's scaling fast.
  • Relocation and Visa Support: We also offer relocation assistance for successful applicants moving to SF, as well as support for visa and green card applications where applicable.
  • Health & Wellness: Comprehensive medical, dental, and vision plans to suit the needs of you and your family.
  • Flexible Spending & Insurance: Access to HSA and FSA accounts, plus life insurance coverage.
  • 401(k) Program: Save for the future with our 401(k) program.
  • Commuter Benefits: We help make getting to and from the office easier and more convenient.
  • Unlimited PTO: So you can take the time you need to recharge, stay healthy, and bring your best self to work.
  • Office Perks: Enjoy a vibrant Downtown San Francisco office with catered lunch five days a week, premium snacks and coffee, an in-building gym, and a dog-friendly environment.

Similar jobs

Senior AI Researcher

Scout AIUnited States· 3 wk ago
RemoteEngineering$200k–$400k/yrapply on job-boards.greenhouse.io