DevOps Engineer
Evlo AI · Washington, DC · Yesterday
RemoteRemoteEngineeringFull-time
About The Role The role owns the reliability, scalability, and security of production infrastructure supporting high-traffic distributed systems and microservices architectures. The team works closely with software engineering and product groups to automate deployments, optimize cloud resources, and maintain zero-downtime reliability standards. Key Responsibilities Design, provision, and manage cloud infrastructure on AWS or GCP using Infrastructure as Code tools like Terraform and TerragruntBuild and maintain robust CI/CD pipelines using GitHub Actions, GitLab CI, or ArgoCD to ensure seamless and safe software deploymentsImplement comprehensive observability stacks using Prometheus, Grafana, Datadog, and ELK to monitor system health and catch anomalies earlyManage Kubernetes clusters in production, handling automated scaling, service meshes, and container security hardeningParticipate in on-call rotations to incident-manage, triage production outages, and conduct blameless root cause analysesEnforce security best practices, identity and access management (IAM) policies, and compliance standards across all cloud environments What We Are Looking For 3–6 years of experience in DevOps engineering, site reliability engineering, or cloud infrastructure managementStrong proficiency with containerization and orchestration technologies, specifically Docker and KubernetesExtensive hands-on experience with Infrastructure as Code (Terraform, CloudFormation) and configuration management toolsDeep working knowledge of Linux system administration, networking fundamentals (TCP/IP, DNS, TLS, VPN), and security protocolsBonus: Experience with service mesh architectures (Istio, Linkerd), FinOps cloud cost optimization, or holding active AWS/Kubernetes certifications