Site Reliability Engineer - AWS - Remote
About the Role
At SitusAMC, we are looking to match your unique experience with one of our careers to help you realize your potential and career growth within the Real Estate Industry. In this role, you will support products recently transitioned from an on-prem data center into AWS Cloud. The position involves strategizing and implementing cloud best practices for newly transitioned products, maintaining operational coverage, and continuously optimizing for efficiency, stability, and reliability. You will enhance automation, scaling, process improvement, metric collection, security, and visibility into product environments.
You will leverage DevOps approaches, including CI/CD processes, and work closely with development teams to ensure secure and manageable migrations into production. You may also embed within product teams to enable enterprise PaaS & SaaS offerings created by the Platform teams.
Responsibilities
- Support and improve production applications running in AWS.
- Monitor system health, availability, latency, and performance.
- Participate in incident response, root cause analysis, and post-incident improvements.
- Improve operational readiness, resiliency, and disaster recovery processes.
- Work with AWS services such as EC2, ECS/EKS, Lambda, S3, RDS, CloudWatch, IAM, VPC, and related cloud services.
- Apply AWS Well-Architected Framework principles across reliability, security, performance, cost optimization, and operational excellence.
- Support cloud-native architecture decisions and infrastructure improvements.
- Improve end-to-end observability across applications, infrastructure, logs, metrics, traces, and alerts.
- Build or enhance dashboards, alerting, and monitoring practices.
- Automate recurring operational tasks using scripting, CI/CD, or infrastructure-as-code tools.
- Support applications hosted both on-premises and in AWS.
- Assist with application modernization and migration from on-prem environments to AWS.
- Partner with application teams to identify performance bottlenecks and reliability risks.
- Work closely with Development, Business, Platform, and Infrastructure teams.
- Participate in Agile delivery using Scrum and Kanban boards.
- Serve as a technical point of contact for product infrastructure and reliability needs.
Requirements
- Bachelor’s degree or equivalent combination of education and experience.
- 5+ years of industry and/or relevant experience, typically with 1+ years in an Associate level role or equivalent.
- Cumulative experience in DevOps – with 70% Ops and 30% Development efforts preferred.
- Strong experience with Containerization, Kubernetes, EKS, and CI/CD pipelines using GitOps Methodology (e.g., ArgoCD, FluxCD).
- Experience with Git for version control; Azure DevOps preferred.
- Experience with monitoring tools (especially CloudWatch alerts), troubleshooting, and serving as an escalation point.
- Implement best practices for configuring, tuning, and securing APIs/Microservices to achieve optimal performance and uptime.
- Utilize tools such as curl and wget to perform diagnostics, analyze network performance, and troubleshoot connectivity issues.
- Deep knowledge of HTTP concepts and protocols to analyze and optimize web application performance.
- Strong experience with Terraform and best practices in Infrastructure-as-Code (IaC).
- Strong scripting and automation skills (Python, Bash, and PowerShell).
- Expertise in AWS Transfer Family, managing EC2 instances and AMIs, deploying applications using AWS Elastic Beanstalk, and configuring/optimizing Load Balancers.
- Proficiency in SQL and SQL administration, managing database/schema/catalog configurations, users/logins, synonyms, and performance troubleshooting in RDS.
- Implement and maintain basic database hygiene concepts, ensuring data integrity and efficient data storage.
- Experience with service mesh technologies (e.g., Linkerd, Istio) is a plus.
- Manage and optimize cloud-based infrastructure in AWS, ensuring high availability and performance.
- Experience with container orchestration tools (Kubernetes, Docker) is required.
- Strong Network Management skills.
- Excellent written and verbal communication skills.
- Experience working in Azure DevOps, including managing and processing tickets, CI/CD pipelines, release management, and infrastructure as code.
- Flexible work hours.
Pay
The annual full-time base salary range for this role is $110,000.00 - $140,000.00. Specific compensation is determined through interviews and a review of relevant education, experience, training, skills, geographic location, and alignment with market data. Certain positions may be eligible to receive a discretionary bonus as determined by bonus program guidelines and SitusAMC Senior Management approval.
Benefits
- Paid Time Off (PTO) and paid holidays, as set forth in program policies.
- Eligibility to participate in various benefit plans, including medical, dental, vision, life, and disability insurance.
- 401K retirement plan participation, in accordance with the terms of the applicable plans.