Jobs · Engineering · Texas

Site Reliability Engineer

Thales · Austin, TX · 2 wk ago
HybridEngineeringFull-time

Location: Austin - Arboretum Plaza, United States (hybrid role in Austin, TX).

About the role

We are looking for an experienced Site Reliability Engineer (SRE) to work with our North American Team. Your responsibility will be to help design and build tools and infrastructure that support our teams and customers. These tools must scale with our growing platform and customer base, driving operational excellence for the company’s globally distributed network. The person in this role will have input in decisions that significantly impact our infrastructure and how we serve our customers.

As an SRE in the ICO organization, you will work with your team to solve problems, support, and optimize the infrastructure programmatically. You will improve the overall availability, reliability, performance, and security of the infrastructure under control.

Responsibilities

  • Work with other teams across the company to build and deploy infrastructure that supports our platforms.
  • Apply SRE core tenets of measurement (SLI/SLO/SLA), eliminate toil, and reliability modeling.
  • Establish metrics for data-driven decisions to increase availability, reliability, and velocity.
  • Build, maintain, and evolve SLO and SLI network/system/application baselines.
  • Assist with go/no-go preplanning, verification/validation, and review of existing and new products/services.
  • Proactively analyze data and test the integrity of networks/systems to ensure production applications and services operate optimally.
  • Work with internal customers to troubleshoot and resolve business-affecting issues.
  • Participate in escalations, incident response, RCAs, and postmortems.
  • Participate in 24x7 on-call rotation.

Requirements

  • At least 5 years of professional experience within a cloud/web/CDN scale infrastructure.
  • Experience with Python and Go. C/C++ is a plus.
  • Expert knowledge of Linux systems, network programming, and protocols (TCP, UDP, DNS, TLS/SSL, HTTP).
  • Experience with BGP and Anycast routing is a plus.
  • Experience with DevOps principles and concepts such as Infrastructure as Code (Ansible/SaltStack), CI/CD (GitLab, Jenkins, Git), monitoring, and visualization (Prometheus, Grafana).
  • Experience with big data technologies such as NoSQL/RDBMS, Redis, ElasticSearch, Kafka.
  • Experience with containers and container management (Docker, Kubernetes).
  • Experience analyzing and building data telemetry, modeling, pipelines, and UI visualization.
  • Experience developing software, troubleshooting, and monitoring large-scale distributed systems.
  • Implement software engineering best practices/standards and software development life cycle.
  • Working knowledge and experience with Agile software development methodologies.
  • Outstanding collaboration, communication, and documentation skills with a proven ability to work cross-functionally.
  • BS/MS in computer science, engineering, or a related technical discipline or equivalent experience.

Schedule

  • Monday to Friday with on-call work over some weekends.
  • Domestic and international travel required occasionally (up to 5 times a year) for department conferences, team meetings, or group working sessions.
  • Hybrid schedule: required to work from the registered Thales/Imperva office a few times a week.

Pay

Total Target Compensation (TTC) market range for this position is between $112,105.50 - $190,664.00 USD annually, inclusive of base salary and variable compensation target. This range reflects how companies in a similar industry and geographic region generally pay for similar jobs and may vary based on career path history, competencies, skills, performance, and internal equity.

Benefits

  • Elective Health, Dental, Vision, FSA/HSA.
  • Voluntary Life and AD&D, Whole Group Life with LTC, Critical Illness, Hospital Indemnity, Accident Insurance.
  • Legal Plan, Identity Theft, and Pet Insurance.
  • Retirement Savings Plan after 30 days of employment with company contribution and match (no vesting period).
  • Company-paid holidays and Paid Time Off.
  • Company-provided Life Insurance, AD&D, Disability, Employee Assistance Plan, and Well-being Program.

Similar jobs

Site Reliability Engineer

Cracker BarrelTennessee, United States· 1 mo ago
Engineeringapply on cbrlgroup.wd503.myworkdayjobs.com