Principal Site Reliability Engineer, Platform
About the role
For our client, we are seeking a Principal Site Reliability Engineer, Platform to join the team of a leader in the Enterprise Technology space. This role will own critical platform infrastructure that supports highly available, scalable software delivery and operations. The principal engineer will help shape the reliability, performance, and automation strategy for a Kubernetes-based environment. This position partners closely with product and engineering teams to launch new capabilities and strengthen the underlying systems that support them. The work will also include backend development, tooling, and continuous improvement efforts that advance uptime and platform effectiveness.
Responsibilities
- Architect and own critical infrastructure
- Build and maintain a Kubernetes-based platform
- Develop backend services in Golang for autonomous systems
- Collaborate with product teams to launch new offerings
- Enhance high availability infrastructure and maintain uptime metrics
- Create tooling to support development and platform operations
- Conduct performance analysis and implement improvements
Qualifications
- 8+ years of experience in infrastructure for high-availability applications
- 6+ years of experience with public cloud solutions
- Expertise in Kubernetes and Terraform
- Strong background in software design methodologies
- Proficiency in Golang, Python, JavaScript, or Rust
- Experience with CI/CD tools like GitHub Actions and ArgoCD
Benefits
- Remote work flexibility within the U.S.
- Possibility of visa sponsorship
- Performance bonuses based on annual evaluations
- Comprehensive benefits package including health and wellness programs
- Focus on diversity, equity, and inclusion initiatives
Location: Remote - US based candidates only, no visa sponsorship available
Compensation: $174,000 – $305,000 annually