Jobs · Engineering · California

Senior Backend Platform Engineer

Thomas To · Santa Clara, CA · Yesterday
EngineeringFull-time

What You'll Be Doing

  • Design, build, and operate production backend services and infrastructure in Go
  • Develop platform capabilities for provisioning, managing, and executing workloads across cloud environments
  • Build reliable control planes, APIs, schedulers, and infrastructure automation
  • Work deeply with Linux, Kubernetes, containers, networking, and public cloud infrastructure
  • Own systems throughout their lifecycle, including architecture, implementation, deployment, observability, incident response, and continuous improvement
  • Solve distributed-systems challenges involving state, concurrency, multi-tenancy, workload isolation, failure recovery, and scalability
  • Build infrastructure and platform primitives used by other engineers and developer-facing products, and establish best practices for system design, code quality, testing, reliability, and production operations
  • Collaborate across engineering and product teams to translate complex infrastructure requirements into simple, dependable developer experiences

What We Need To See

  • B.S. degree or equivalent experience
  • 8+ years of relevant software engineering experience, with flexibility for exceptional candidates
  • Strong professional experience developing production systems in Go, Linux systems knowledge and the ability to debug across system layers
  • Strong networking fundamentals, including TCP/IP, DNS, routing, proxies, VPNs, and load balancing
  • Hands-on experience with Kubernetes and containerized workloads, and building infrastructure on AWS, GCP, or Azure
  • Backend or platform engineering experience with production systems and strong distributed-systems fundamentals, including consistency, fault tolerance, concurrency, and failure handling
  • Experience building infrastructure, developer platforms, cloud services, or shared systems that other engineers depend on
  • A track record of owning reliability and operational outcomes in addition to feature delivery

Ways To Stand Out From The Crowd

  • Experience with Temporal or another durable workflow orchestration system
  • Experience designing multi-tenant platforms, control planes, or schedulers
  • Knowledge of VM lifecycle management, remote execution environments, or sandbox and isolation technologies
  • Experience with observability, reliability engineering, capacity planning, or infrastructure automation and building AI agent platforms or developer execution environments
  • Experience building and operating GPU infrastructure, including GPU provisioning, scheduling, orchestration, or workload management

Pay

Base salary range is 184,000 USD - 287,500 USD for Level 4, and 224,000 USD - 356,500 USD for Level 5. You will also be eligible for equity and benefits.

Similar jobs