Senior Cloud/ML Ops Engineer
Ivo · San Francisco, CA · 2 wk ago
On-siteEngineering$250k–$325k/yrFull-time
The Role
Why? Infrastructure Engineers build the foundation for Ivo’s entire platform. Customers are cagey about their contracts, so each customer gets their isolated environment with containers, database, VPC, etc. Things break. Regions go down. Cloud and LLM providers have “incidents.” Customers still expect us to hit our SLAs.
What? We’re looking for an Cloud/MLOps Engineer as part of Infrastructure team to:
- Own and evolve our Kubernetes platform across AWS/GCP/Azure
- Design strategies to isolate ML vs API workloads while optimizing for cost, performance, and reliability
- Implement security and compliance controls at the platform layer with RBAC, workload identity, secrets management while preserving data isolation aligned with residency requirements and auditability for enterprise customers
- Partner with SRE + ML teams to ensure SLOs are realistic and enforceable and models are deployed in production environments
Who
We need someone who has:
- Deep, hands-on experience with Kubernetes in production (you’ve debugged it at 2am, not just deployed to it)
- Strong experience with infrastructure as code (Pulumi, Terraform, etc.)
- A strong understanding of cluster architecture, scheduling, networking, storage primitives and failure modes in distributed systems
- Experience managing multi-cluster or multi-region setups with Github CI/CD
Qualifications
- Experience working in a startup environment is preferred but not required
- Excited about the adventure of building a company!