Jobs · Engineering · New York

Engineering Manager, Artifacts & Registry - W&B

Weights & Biases · New York, NY · 2 wk ago
Engineering$165k–$242k/yrFull-time

About the Role

You'll join the ML Workflows organization as the Engineering Manager for Artifacts & Registry, leading a team of six engineers. Registry is the organization-level source of truth for what a company has approved for production. It provides access control, protected aliases, lineage, and automation hooks, and it is where a model version goes once it moves out of research. Pinterest, Capital One, Canva, Recursion, and Dropbox all run on it, and for some of them it sits inside a production serving path.

Artifacts is the layer underneath it: content-addressable storage, deduplication, lineage, and retention for datasets, model weights, checkpoints, and evaluation outputs, across S3, GCS, Azure, and CoreWeave AI Object Storage. Every object logged to W&B is an artifact.

The team's mandate is also weighted heavily toward closing the gap between CoreWeave's infrastructure and the ML experimentation that happens in W&B. That includes extending the Registry beyond a record of decisions, so that promoting a model version can put it into serving on CoreWeave compute with lineage and access control carried through, and bringing compute and experimentation into a single unified experience rather than separate tools. This work is done in partnership with foundational infrastructure teams at CoreWeave and product teams across the W&B ecosystem.

The current scope is well defined and load bearing, with enterprise customers relying on it daily. What is open is what comes next, where there is significant room for new products, new roadmaps, and deeper integration between CoreWeave and W&B. The role therefore calls for a product-first mindset alongside strong operational ownership.

Responsibilities

  • Own the full lifecycle, from roadmap shaping through delivery and production reliability.
  • Partner with product management on prioritization and unblock engineers across team boundaries.
  • Maintain the reliability bar while the team ships new capability, ensuring customer data and uptime are protected.
  • Engage in technical architecture conversations, including relational data modeling, query performance, object storage, permission models, and data lifecycle management.
  • Grow engineers into autonomous leaders and technical decision-makers, fostering an AI-forward mindset.
  • Champion operational accountability, making hard prioritization calls and building a team that operates with autonomy and urgency.

Requirements

  • 3+ years of engineering management experience shipping platform, data, or infrastructure products.
  • 5+ years of software engineering experience prior to management, with backend systems depth.
  • Technical fluency in Go (or a comparable statically typed backend language), relational data modeling, and distributed backend systems.
  • Experience managing teams that own production systems where the durability and correctness of customer data is the primary constraint.
  • Demonstrated ability to balance product delivery with technical health, covering technical debt, reliability, and on-call.
  • Experience partnering with product managers to shape a roadmap, including making the case for platform investment with evidence.
  • Track record of growing engineers across levels, from mid-level through senior.
  • Strong communication skills with both engineering and non-engineering stakeholders.
  • Comfortable operating in ambiguity, including product areas where the right level of investment is an open question.

Preferred Qualifications

  • Experience with ML infrastructure, ML platforms, model registries, or developer tools for data scientists.
  • Background in large-scale object storage, content-addressable storage, deduplication, or garbage collection over large object graphs.
  • Depth in query performance and schema migrations against large production databases, and experience with a columnar or OLAP store such as ClickHouse alongside a transactional system of record.
  • Familiarity with multi-tenant permission models, RBAC, and enterprise access control requirements.
  • Experience taking a mature, widely-adopted product into a new strategic phase rather than only building from zero.
  • Experience with post-acquisition platform integration.
  • Founding experience or early-stage background where you took a product from idea to the hands of users.

Skills

  • You're interested in owning a system other products depend on, and in extending it rather than only keeping it running.
  • You think reliability is part of the product, and you can make that case to product partners and leadership with data.
  • You care about developer experience and understand that infrastructure teams succeed when their users actually love the tools.
  • You've managed through organizational complexity before and are comfortable navigating cross-team dependencies and shifting priorities.

Benefits

  • Medical, dental, and vision insurance - 100% paid for by CoreWeave.
  • Company-paid life insurance.
  • Voluntary supplemental life insurance.
  • Short and long-term disability insurance.
  • Flexible Spending Account (FSA) and Health Savings Account (HSA).
  • Tuition reimbursement.
  • Employee Stock Purchase Program (ESPP).
  • Mental wellness benefits through Spring Health.
  • Family-forming support provided by Carrot.
  • Paid parental leave.
  • Flexible, full-service childcare support with Kinside.
  • 401(k) with a generous employer match.
  • Flexible PTO.
  • Catered lunch each day in offices and data center locations.
  • A casual work environment and a work culture focused on innovative disruption.

Pay

The base salary range for this role is $165,000 to $242,000. The starting salary will be determined based on job-related knowledge, skills, experience, and market location. In addition to base salary, total rewards include a discretionary bonus, equity awards, and a comprehensive benefits program.

Similar jobs