Software Engineer, DevOps
About the Role
We are seeking an experienced DevOps Engineer to join our growing team and play a pivotal role in designing and building our platform and infrastructure as we continue to scale our product and user base. As part of our team, you will work in a dynamic, fast-paced environment to ensure the reliability, scalability, and performance of our systems, with a focus on service architecture, deployment, query optimization, distributed systems, data and machine learning infrastructure, and security and authentication.
Responsibilities
- Partner with product teams to architect, design, and build the foundational infrastructure for our products.
- Design, develop, and deploy highly available and scalable Multi-tenant SaaS solutions on public cloud networks like AWS, Azure, and GCP.
- Leverage technologies such as Kubernetes, Helm, Terraform, and Istio to achieve infrastructure resilience.
- Drive the automation of infrastructure tasks, from provisioning to configuration management and deployment, using tools like Terraform, Ansible, and Kubernetes.
- Collaborate with the software development team to refine CI/CD pipelines (e.g., GitHub Actions, Cloud Build), enhance service interfaces, and improve the overall developer experience.
- Architect and implement advanced observability solutions using tools like Prometheus and Grafana.
- Ensure real-time alerting and error tracking with Sentry and Pagerduty to maintain system health and performance.
- Deploy comprehensive testing frameworks, including Selenium for end-to-end testing, to ensure robust integration and system testing.
- Regularly monitor system health, analyze performance metrics, and recommend enhancements, including optimizing database queries for peak performance.
Qualifications
- Bachelor's or Master's degree in Computer Science or a related field.
- 3+ years of experience in Infrastructure engineering or a similar role.
- Excellent problem-solving skills and the ability to work under pressure in a fast-paced environment.
- Ability to work independently and as part of a team.
- Experience working with global teams.
Nice to Have
- ML/Ops experience.
- Experience with Postgres query optimization and related performance improvement techniques.
- Experience with event-driven data and machine learning infrastructure, including streaming pipelines, database systems, and model training.
- Experience with air-gapped cloud environments or private clouds.
- Experience administering complex deployments on Azure, especially AKS.
Pay
For California-based candidates, the standard base salary for this position is $135,000-$225,000 annually. Compensation offered will be determined by factors such as location, level, job-related knowledge, skills, and experience. Certain roles may be eligible for variable compensation, equity, and benefits.