Jobs · Information Technology · Georgia

Site Reliability Engineer

Autodesk · Atlanta, GA · 2 wk ago
HybridInformation Technology$117k–$209k/yrFull-time

Responsibilities

  • Architect and implement hosting solutions for highly dynamic SaaS web applications, ensuring reliability and performance at scale
  • Design, implement, and maintain Infrastructure-as-Code solutions to support scalable, reliable, and secure global environments
  • Implement security and compliance with best practices across infrastructure and applications, including hardening, enforcing least privileges
  • Use modern administration tools like Docker, Terraform, AWS CloudFormation/CDK to manage and deploy containers and virtual machines
  • Collaborate with development, testing, and documentation teams during the product development cycle to ensure quality control
  • Automate processes and integrate new technologies as needed
  • Define and monitor Service Level Objectives (SLOs), Service Level Indicators (SLIs), and manage error budgets to ensure reliability goals are met
  • Work with stakeholders to align technical strategy with business requirements
  • Participate in on-call support and incident management, ensuring timely resolution and clear communication
  • Conduct post-incident blameless postmortems to identify root causes and drive continuous improvement

Requirements

  • 5+ years DevOps/SRE experience with cloud-based applications
  • Advanced hands-on experience Linux administration skills, including monitoring, troubleshooting, reliability and security
  • Must have U.S. citizenship or have U.S. lawful permanent residency
  • Experience managing large-scale cloud infrastructure (AWS preferred)
  • Strong scripting abilities (eg. Bash, Python, Perl, etc.)
  • Expert-level knowledge of AWS services (EC2, ECS, EKS, Lambda, ELB, S3, IAM, VPC, Dynamo DB, RDS, etc.)
  • Hands-on experience with Docker, Kubernetes and container technologies
  • Proficiency with infrastructure-as-code tools (Terraform, CloudFormation)
  • Experience with CI/CD tools. (Jenkins, Artifactory, GIT, etc.)
  • Skilled in log analysis and monitoring tools (CloudWatch, Splunk, Dynatrace, New Relic, Grafana)
  • Experience with relational databases (MySQL, PostgreSQL, MSSQL), and SQL
  • Excellent problem-solving skills and ability to work independently
  • Excellent written and verbal communication skills

Qualifications

  • Bachelor's degree in computer science or related field

Skills

  • DevOps/SRE experience
  • Linux administration
  • Cloud infrastructure management (AWS preferred)
  • Scripting (Bash, Python, Perl, etc.)
  • AWS services (EC2, ECS, EKS, Lambda, ELB, S3, IAM, VPC, Dynamo DB, RDS, etc.)
  • Docker, Kubernetes, container technologies
  • Infrastructure-as-code tools (Terraform, CloudFormation)
  • CI/CD tools (Jenkins, Artifactory, GIT, etc.)
  • Log analysis and monitoring tools (CloudWatch, Splunk, Dynatrace, New Relic, Grafana)
  • Relational databases (MySQL, PostgreSQL, MSSQL)
  • Problem-solving skills
  • Communication skills

Benefits

Salary range: $117,000 - $209,330
Benefits: Health and financial benefits, time away, everyday wellness, stock grants, and a comprehensive benefits package.

Pay

Starting base salary between $117,000 and $209,330.

Schedule

Not specified.

Similar jobs

Site Reliability Engineer

QualityAIIllinois, United States· 2 days ago
Quality Assurance$110k–$120k/yrapply on careers.quality-ai.com