AI Research Scientist, New Grad – Agents & Reinforcement Learning
SLAS (Society for Laboratory Automation and Screening) · Bellevue, WA · 1 wk ago
EngineeringOther
Snowflake is seeking an AI Research Scientist (New Grad) for its AI Research team. This role is at the intersection of agentic AI and reinforcement learning, where your research will directly shape how enterprises leverage intelligent automation.
About the role
- Design and develop agentic frameworks powered by recursive self-improvement loops, enabling AI systems that iteratively refine their own capabilities and strategies
- Build and evaluate auto research agents — systems capable of autonomously formulating hypotheses, executing experiments, and synthesizing findings
- Develop coding agents that understand, generate, and debug code across complex, multi-step programming tasks
- Conduct research in reinforcement learning with a focus on RLHF, DPO, and PPO as mechanisms for aligning and improving agentic behaviors
- Contribute to multi-agent systems where specialized agents collaborate, negotiate, and self-organize to solve enterprise-scale problems
- Develop and curate training data pipelines — both synthetic and human-annotated — to support novel agentic and RL research domains
Responsibilities
- Design and develop agentic frameworks powered by recursive self-improvement loops, enabling AI systems that iteratively refine their own capabilities and strategies
- Build and evaluate auto research agents — systems capable of autonomously formulating hypotheses, executing experiments, and synthesizing findings
- Develop coding agents that understand, generate, and debug code across complex, multi-step programming tasks
- Conduct research in reinforcement learning with a focus on RLHF, DPO, and PPO as mechanisms for aligning and improving agentic behaviors
- Contribute to multi-agent systems where specialized agents collaborate, negotiate, and self-organize to solve enterprise-scale problems
- Develop and curate training data pipelines — both synthetic and human-annotated — to support novel agentic and RL research domains
Requirements
- PhD in Computer Science, Machine Learning, Artificial Intelligence, or a closely related field (completing or recently completed; or equivalent research experience)
- Foundational expertise in reinforcement learning algorithms, including RLHF, DPO, PPO, or multi-agent systems
- Research experience in LLM post-training, fine-tuning, or reasoning model development
- Demonstrated ability to implement and experiment with agentic architectures — including tool-use, planning, and self-correction loops
- Proficiency in Python and at least one deep learning framework (PyTorch or JAX strongly preferred)
- Strong mathematical and analytical foundation — comfortable working at the intersection of theory and empirical research
Qualifications
- At least one first-author or co-authored publication or preprint in a relevant AI/ML area
- Hands-on experience building or evaluating coding agents or auto research agents
- Familiarity with recursive self-improvement frameworks or automated AI scientist paradigms
- Experience with large-scale distributed training or efficient training paradigms
- Background in mathematical reasoning, structured decision-making, or program synthesis
- Exposure to domain-specific AI applications in healthcare, finance, or enterprise workflows
Skills
- Python
- Deep learning frameworks (PyTorch or JAX preferred)
- Reinforcement learning algorithms (RLHF, DPO, PPO)
- Multi-agent systems
- Agentic architectures
- Recursive self-improvement frameworks
Benefits
- Opportunity to contribute to systems that actually run in production
- Access to large-scale compute, real-world data challenges, and the shortest possible path from research idea to product impact
- Scaling the team to help enable and accelerate Snowflake's growth
- Collaboration with researchers and engineers building the next generation of agentic infrastructure
Pay
Please visit the job posting on the Snowflake Careers Site for salary and benefits information: careers.snowflake.com
Schedule
Full-time position