Cloud Infrastructure Site Reliability Engineer (SRE)
EPITEC · Berkeley Heights, NJ · 2 wk ago
Information Technology$60–$94/hrContract
Location: Alpharetta, GA or Berkeley Heights, NJ (5 days onsite)
Duration: 12 months plus extensions
Employment type: W2 only
Eligibility: Green Card or Citizens only
Pay Rate Range: $60–94/hr
About the role
As a Cloud Infrastructure Site Reliability Engineer (SRE) with expertise in multiple public-cloud platforms, you will operate infrastructure solutions following the principles and practices pioneered by Google’s SRE model. Your work will ensure cloud services meet uptime, reliability, and performance targets, and you will drive automation and continuous improvement across production environments. This role involves collaborating with cross-functional teams to enhance cloud reliability posture and streamline processes through automation.
Requirements
- Bachelor’s degree in Computer Science, Engineering, or a related technical field, or equivalent practical experience
- 3+ years of experience in software development with proficiency in at least one programming language (e.g., Python, Go, Java, C++)
- Experience administrating cloud platforms (AWS, GCP, Azure), including networking, security, containerization, storage, data management, and serverless technologies
- Solid understanding of Linux systems, networking fundamentals, virtualized and distributed systems, file systems, system processes, and configurations
- Deep understanding of observability (monitoring, alerting, and logging) tools in cloud environments; ability to set up and maintain monitoring dashboards, alerts, and logs
- Familiarity with Continuous Integration/Continuous Deployment (CI/CD) tools for automated testing, deployments, provisioning, and observability
- Ability to manage and respond to incidents, perform root-cause analysis, and implement post-mortem reviews
- Understanding of setting, monitoring, and maintaining Service-Level Objectives (SLOs) and Service-Level Agreements (SLAs) for system reliability
- Experience with Terraform and Dynatrace
Nice-to-have qualifications
- Experience working with enterprise-scale financial services or other regulated industries
- 5+ years of experience in SRE, DevOps, infrastructure, or cloud engineering roles, preferably supporting large-scale, distributed systems
- Excellent problem-solving, troubleshooting, and communication skills
- Experience leading technical projects or mentoring junior engineers
- Relevant certifications: Certified Engineer, DevOps, SRE, CSREF