Senior DevOps / Cloud Engineer
ECS · Arlington, VA · 3 days ago
$140k–$200k/yrFull-time
ECS is seeking a Senior DevOps/Cloud Engineer to work in our Arlington, VA (hybrid) office. We are looking for a talented engineer who is passionate about building, automating, and operating modern cloud environments.
About the role
This is a hands-on engineering role focused on designing, building, automating, and operating cloud platforms and delivery pipelines. The ideal candidate will be comfortable writing code and automation, troubleshooting complex infrastructure problems, implementing cloud-native solutions, and partnering directly with software, security, and infrastructure teams.
Responsibilities
- Design, build, operate, and improve scalable, reliable, and secure cloud infrastructure.
- Develop and maintain infrastructure-as-code using technologies such as Terraform, CloudFormation, Bicep, or equivalent tools.
- Build, maintain, and improve CI/CD pipelines supporting application, infrastructure, and configuration deployments.
- Develop automation, tooling, scripts, and integrations using languages such as Python, Go, PowerShell, or Bash.
- Deploy and operate containerized applications using Docker, Kubernetes, and cloud-native container platforms.
- Design and implement cloud architectures using services across AWS, Azure, and/or Google Cloud Platform.
- Automate provisioning, configuration management, deployment, monitoring, scaling, and operational workflows.
- Support the development and operation of microservices, APIs, distributed applications, and cloud-native platforms.
- Implement monitoring, logging, alerting, observability, and performance-management capabilities across applications and infrastructure.
- Troubleshoot complex issues involving Linux systems, networking, containers, cloud services, application deployments, and automation pipelines.
- Improve system reliability, availability, resiliency, performance, and operational efficiency through automation and engineering.
- Implement and maintain secure configuration, identity and access controls, secrets management, certificates, and cloud security best practices.
- Partner with software engineers to improve development workflows, deployment processes, and application operability.
- Partner with security engineers to integrate security controls into cloud infrastructure, CI/CD pipelines, and development workflows.
- Identify repetitive operational processes and replace them with scalable automation.
- Evaluate new cloud, DevOps, platform engineering, and automation technologies and determine where they can improve the environment.
- Provide technical leadership and mentorship while remaining actively involved in implementation and troubleshooting.
Requirements
- Bachelor’s degree in engineering, computer science, information technology, or another technical discipline, or an associate degree with equivalent years of relevant technical experience.
- At least 8 years of overall technical experience, including significant hands-on experience in several of the following areas:
- DevOps Engineering
- Cloud Engineering
- Platform Engineering
- Unix/Linux Administration
- Infrastructure Automation
- CI/CD Engineering
- Networking and Virtual Infrastructure
- Containerization and Kubernetes
- Application Deployment and Operations
- Software Development or Scripting
- Strong scripting or development experience in at least one language such as Python, Go, PowerShell, Bash, Perl, or similar.
- Demonstrated experience designing, building, and operating production infrastructure using automation and infrastructure-as-code.
- Demonstrated ability to independently troubleshoot complex infrastructure, cloud, networking, application, and deployment issues.
- Ability to work directly with software, security, and infrastructure teams to translate technical requirements into reliable and maintainable engineering solutions.
- Must successfully complete a stringent Background Investigation and obtain the required Government Security Clearance.
Skills
- Strong understanding of Git and modern source-control workflows.
- Strong experience with Python, Golang, C#, Java, or another general-purpose programming language.
- Strong experience with Bash and/or PowerShell.
- Deep knowledge of Linux administration, troubleshooting, performance tuning, and core system services.
- Strong understanding of TCP/IP networking, DNS, HTTP/S, load balancing, firewalls, routing, proxies, and network troubleshooting.
- Hands-on experience with AWS, Azure, and/or Google Cloud Platform.
- Experience with Terraform, CloudFormation, Bicep, Pulumi, or equivalent infrastructure-as-code technologies.
- Experience with Docker, Kubernetes, Helm, and container orchestration platforms.
- Experience designing and maintaining CI/CD pipelines using platforms such as GitHub Actions, GitLab CI/CD, Jenkins, Azure DevOps, or equivalent tools.
- Experience with configuration-management technologies such as Ansible, Puppet, Chef, or equivalent tools.
- Experience implementing monitoring, logging, alerting, and observability using platforms such as Prometheus, Grafana, OpenTelemetry, Splunk, Elastic, CloudWatch, Azure Monitor, or equivalent technologies.
- Experience with cloud identity and access management, secrets management, encryption, certificates, and secure cloud configuration.
- Familiarity with Site Reliability Engineering (SRE) concepts such as service-level objectives, availability, resiliency, capacity management, and incident response.
- Experience supporting highly available, fault-tolerant, and distributed systems.
- Experience integrating automated testing, security scanning, policy checks, and deployment controls into CI/CD workflows.
- Familiarity with relational and NoSQL databases, messaging systems, caching platforms, and other common cloud application services.
Pay
Salary Range: $140,000-$200,000