Cloud Operations Engineer
Alliance of Professionals & Consultants, Inc. (APC) · Pittsburgh, PA · Yesterday
EngineeringContract
About the role
Manages design, development, implementation, and maintenance of cloud computing infrastructure including gateways to remote services and endpoint systems.
Responsibilities
- Provision, configure, and maintain AWS resources such as EC2 instances, VPCs, S3 buckets, RDS databases, and other services.
- Maintain accurate documentation of cloud architecture, configurations, and operation procedures.
- Monitor system performance, availability, and health using Amazon CloudWatch and other monitoring tools.
- Create and manage IAM users, roles, and policies to enforce least-privilege access and maintain security compliance.
- Troubleshoot and resolve infrastructure, networking, and service-related incidents within AWS environments.
- Manage networking components including VPCs, transit gateways, subnets, route tables, load balancers, and Route 53.
- Implement logging, auditing, and governance controls using AWS Config, CloudTrail, and organizational policies.
- Configure auto scaling, load balancing, and multi-AZ deployments to ensure high availability and scalability.
- Automate infrastructure deployment and updates using Infrastructure as Code tools such as Terraform and AWS CloudFormation.
- Deploy, manage, and maintain containerized applications using Amazon Elastic Kubernetes Service (EKS), including cluster configuration, node management, and workload scaling.
- Manage backups, snapshots, and disaster recovery solutions to ensure business continuity.
- Configure patch management and system updates using AWS Systems Manager to maintain secure environments.
- Monitor and optimize AWS costs using tools such as AWS Cost Explorer, implementing rightsizing and savings plans as needed.
- Configure and maintain security controls including security groups, NACLs, encryption, AWS Shield, and GuardDuty.
Requirements
- Extensive experience deploying and managing infrastructure within Amazon Web Services (AWS) environments.
- Advanced proficiency in Infrastructure as Code (IaC) using Terraform, including module development, state management, and multi-environment deployments.
- Strong knowledge of AWS core services including EC2, S3, RDS, EBS, Lambda, and Elastic Load Balancing.
- Deep understanding of AWS Identity and Access Management (IAM), AWS Organizations, and role-based access controls.
- Familiarity with hybrid cloud integrations between AWS and on-premises environments.
- Experience designing and managing AWS backup, disaster recovery, and business continuity solutions (e.g., AWS Backup, cross-region replication).
- Experience managing containerized workloads using Amazon EKS and ECS, including cluster configuration and operational support.
- Strong working knowledge of AWS networking including VPCs, subnets, route tables, NAT gateways, Transit Gateway, Direct Connect, Route 53, and security groups.
- Excellent written and verbal communication skills, including the ability to create technical documentation, architecture diagrams, and operational runbooks.
- Commitment to operational excellence and adherence to AWS Well-Architected Framework principles.
- Strong analytical and problem-solving skills in complex, automated cloud environments.
- Adaptability and willingness to work in DevOps-driven environments with rapidly evolving priorities.
- High attention to detail in infrastructure configuration, security controls, and cost optimization.
- Prioritize multiple infrastructure initiatives and operational tasks while meeting deadlines.
- Sound judgment with the ability to make timely, risk-aware decisions in highly available AWS environments.
Qualifications
- Extensive experience deploying and managing infrastructure within Amazon Web Services (AWS) environments.
- Advanced proficiency in Infrastructure as Code (IaC) using Terraform, including module development, state management, and multi-environment deployments.
- Strong knowledge of AWS core services including EC2, S3, RDS, EBS, Lambda, and Elastic Load Balancing.
- Deep understanding of AWS Identity and Access Management (IAM), AWS Organizations, and role-based access controls.
- Familiarity with hybrid cloud integrations between AWS and on-premises environments.
- Experience designing and managing AWS backup, disaster recovery, and business continuity solutions (e.g., AWS Backup, cross-region replication).
- Experience managing containerized workloads using Amazon EKS and ECS, including cluster configuration and operational support.
- Strong working knowledge of AWS networking including VPCs, subnets, route tables, NAT gateways, Transit Gateway, Direct Connect, Route 53, and security groups.
- Excellent written and verbal communication skills, including the ability to create technical documentation, architecture diagrams, and operational runbooks.
- Commitment to operational excellence and adherence to AWS Well-Architected Framework principles.
- Strong analytical and problem-solving skills in complex, automated cloud environments.
- Adaptability and willingness to work in DevOps-driven environments with rapidly evolving priorities.
- High attention to detail in infrastructure configuration, security controls, and cost optimization.
- Prioritize multiple infrastructure initiatives and operational tasks while meeting deadlines.
- Sound judgment with the ability to make timely, risk-aware decisions in highly available AWS environments.
Skills
- Extensive experience deploying and managing infrastructure within Amazon Web Services (AWS) environments.
- Advanced proficiency in Infrastructure as Code (IaC) using Terraform, including module development, state management, and multi-environment deployments.
- Strong knowledge of AWS core services including EC2, S3, RDS, EBS, Lambda, and Elastic Load Balancing.
- Deep understanding of AWS Identity and Access Management (IAM), AWS Organizations, and role-based access controls.
- Familiarity with hybrid cloud integrations between AWS and on-premises environments.
- Experience designing and managing AWS backup, disaster recovery, and business continuity solutions (e.g., AWS Backup, cross-region replication).
- Experience managing containerized workloads using Amazon EKS and ECS, including cluster configuration and operational support.
- Strong working knowledge of AWS networking including VPCs, subnets, route tables, NAT gateways, Transit Gateway, Direct Connect, Route 53, and security groups.
- Excellent written and verbal communication skills, including the ability to create technical documentation, architecture diagrams, and operational runbooks.
- Commitment to operational excellence and adherence to AWS Well-Architected Framework principles.
- Strong analytical and problem-solving skills in complex, automated cloud environments.
- Adaptability and willingness to work in DevOps-driven environments with rapidly evolving priorities.
- High attention to detail in infrastructure configuration, security controls, and cost optimization.
- Prioritize multiple infrastructure initiatives and operational tasks while meeting deadlines.
- Sound judgment with the ability to make timely, risk-aware decisions in highly available AWS environments.
Benefits
Not specified.
Pay
A reasonable estimate of the pay range for this role is $00.00 - $00.00 per hour.
Schedule
Hybrid in Pittsburgh, PA.