Jobs · Science

Senior Research Scientist LLM

techire ai · San Francisco, CA · 1 wk ago
RemoteRemoteScience$500k/yrFull-time

About the role

This team is building the next generation of conversational AI, working alongside an ex-NVIDIA & Meta research leader and one of the co-creators behind well-known open-source models. The company is Series A, experiencing rapid growth with strong commercial traction, and their technology powers hundreds of millions of conversations every month. You'll work on research that directly improves models deployed at real-world scale.

What you'll do

  • Design and run post-training experiments using techniques such as DPO, GRPO, SFT, rejection sampling and related approaches
  • Build and scale reinforcement learning and post-training infrastructure
  • Develop reward models, evaluation frameworks and assessment rubrics that improve model quality and behaviour
  • Work with vendors and data partners to build high-quality post-training datasets and feedback pipelines
  • Own end-to-end projects spanning data collection, modelling, evaluation and infrastructure, identifying where models break down and driving improvements in reasoning, controllability and alignment

What you'll bring

We're looking for someone who's owned significant post-training projects that have delivered meaningful capability improvements. You'll understand the practical challenges behind modern post-training techniques, not just the theory. You've worked across data, modelling, evaluation and RL infrastructure, know how to improve model quality through robust evaluation and preference optimisation, and understand what it takes to make capability improvements repeatable.

Pay

Depends on experience/location. In SF, comp is up to $500,000 base salary, plus stock.

Similar jobs