Staff Software Engineer – Cloud Platform (FortiSIEM)
Fortinet · Santa Clara, CA · Yesterday
HybridEngineering$179k–$219k/yrFull-time
Responsibilities
- Design, build, and maintain cloud platform infrastructure that provisions and manages customer environments.
- Develop automation for deployment, provisioning, upgrades, scaling, backup, and disaster recovery.
- Build and enhance platform services written in Python and other backend technologies that orchestrate cloud resources and manage platform lifecycle.
- Improve infrastructure-as-code using Terraform/OpenTofu and related automation frameworks.
- Enhance monitoring, observability, alerting, and operational tooling across platform and tenant environments.
- Improve platform reliability through high availability, resiliency, and disaster recovery capabilities.
- Design and improve CI/CD pipelines, deployment automation, and rollback mechanisms.
- Troubleshoot complex production issues across cloud infrastructure, backend services, databases, networking, and deployment systems.
- Partner with product, QA, SRE, and security teams to deliver secure and highly available cloud services.
- Drive continuous improvements in operational efficiency, automation, and platform scalability.
- Participate in production support and on-call rotation to ensure service reliability.
Qualifications
- 8+ years of software engineering experience with at least 4 years in cloud platform or infrastructure engineering.
- Strong Python programming skills.
- Experience with Infrastructure-as-Code using Terraform, OpenTofu, or similar technologies.
- Experience building cloud-native services using containers and Kubernetes.
- Experience designing and operating highly available distributed systems.
- Strong Linux, networking, and troubleshooting skills.
- Experience with CI/CD pipelines and deployment automation.
- Excellent problem-solving skills with a systems-thinking mindset.
Preferred Experience
- Experience with one or more cloud platforms such as AWS, OpenStack, Azure, or Google Cloud.
- Experience with ClickHouse or other distributed databases.
- Experience with observability platforms, monitoring, logging, and alerting.
- Experience designing disaster recovery and high availability solutions.
- Experience building SaaS platforms or operating large-scale production services.
- Experience in cybersecurity, SIEM, or cloud security products.