Site Reliability Engineer
Autodesk · Atlanta, GA · 2 wk ago
HybridInformation Technology$117k–$209k/yrFull-time
Responsibilities
- Architect and implement hosting solutions for highly dynamic SaaS web applications, ensuring reliability and performance at scale
- Design, implement, and maintain Infrastructure-as-Code solutions to support scalable, reliable, and secure global environments
- Implement security and compliance with best practices across infrastructure and applications, including hardening, enforcing least privileges
- Use modern administration tools like Docker, Terraform, AWS CloudFormation/CDK to manage and deploy containers and virtual machines
- Collaborate with development, testing, and documentation teams during the product development cycle to ensure quality control
- Automate processes and integrate new technologies as needed
- Define and monitor Service Level Objectives (SLOs), Service Level Indicators (SLIs), and manage error budgets to ensure reliability goals are met
- Work with stakeholders to align technical strategy with business requirements
- Participate in on-call support and incident management, ensuring timely resolution and clear communication
- Conduct post-incident blameless postmortems to identify root causes and drive continuous improvement
Requirements
- 5+ years DevOps/SRE experience with cloud-based applications
- Advanced hands-on experience Linux administration skills, including monitoring, troubleshooting, reliability and security
- Must have U.S. citizenship or have U.S. lawful permanent residency
- Experience managing large-scale cloud infrastructure (AWS preferred)
- Strong scripting abilities (eg. Bash, Python, Perl, etc.)
- Expert-level knowledge of AWS services (EC2, ECS, EKS, Lambda, ELB, S3, IAM, VPC, Dynamo DB, RDS, etc.)
- Hands-on experience with Docker, Kubernetes and container technologies
- Proficiency with infrastructure-as-code tools (Terraform, CloudFormation)
- Experience with CI/CD tools. (Jenkins, Artifactory, GIT, etc.)
- Skilled in log analysis and monitoring tools (CloudWatch, Splunk, Dynatrace, New Relic, Grafana)
- Experience with relational databases (MySQL, PostgreSQL, MSSQL), and SQL
- Excellent problem-solving skills and ability to work independently
- Excellent written and verbal communication skills
Qualifications
- Bachelor's degree in computer science or related field
Skills
- DevOps/SRE experience
- Linux administration
- Cloud infrastructure management (AWS preferred)
- Scripting (Bash, Python, Perl, etc.)
- AWS services (EC2, ECS, EKS, Lambda, ELB, S3, IAM, VPC, Dynamo DB, RDS, etc.)
- Docker, Kubernetes, container technologies
- Infrastructure-as-code tools (Terraform, CloudFormation)
- CI/CD tools (Jenkins, Artifactory, GIT, etc.)
- Log analysis and monitoring tools (CloudWatch, Splunk, Dynatrace, New Relic, Grafana)
- Relational databases (MySQL, PostgreSQL, MSSQL)
- Problem-solving skills
- Communication skills
Benefits
Salary range: $117,000 - $209,330
Benefits: Health and financial benefits, time away, everyday wellness, stock grants, and a comprehensive benefits package.
Pay
Starting base salary between $117,000 and $209,330.
Schedule
Not specified.