Senior Site Reliability Engineer
Boeing · Berkeley, MO · 1 mo ago
Engineering$161k–$217k/yrFull-time
About the role
The Boeing Company is looking for a Senior Site Reliability Engineer to join the Air Dominance Site Reliability Engineering team located in Berkeley, MO.
Responsibilities
- Operate and maintain GitLab, GitLab CI/CD runners, Jira, Confluence, PostgreSQL, and related developer tooling infrastructure
- Lead the deployment and configuration of software development technologies, including build servers, version control systems, CI/CD pipelines, and automated testing frameworks
- Serve as a technical owner for platform reliability, availability, performance, capacity, backup, recovery, and operational readiness
- Lead troubleshooting for complex application, database, runner, pipeline infrastructure, network, storage, and performance issues
- Develop and maintain Infrastructure as Code (IaC), Ansible, and other automation for provisioning, configuration, platform scaling, health checks, reporting, backup validation, and routine operational tasks
- Plan and execute approved changes, including application upgrades, security patches, database maintenance, runner lifecycle activities, and infrastructure updates
- Establish and improve monitoring, alerting, dashboards, SLIs, SLOs, SLAs, KPIs, error budgets, and operational metrics
- Support incident response, root cause analysis, corrective action tracking, and post-incident reviews
- Mentor junior engineers and provide technical guidance on SRE practices, secure administration, automation, and troubleshooting
- Improve runbooks, standard operating procedures, architecture documentation, and disaster recovery procedures
- Evaluate platform risks, capacity trends, recurring incidents, and operational toil, then recommend and implement improvements
- Lead demonstrations, monitor progress, and present technical status to customers and management
- Lead process improvement efforts that help operationally field higher-quality end-to-end system software more frequently
- Participate in after-hours support for urgent or mission-impacting issues as required
Qualifications
- Active Secret U.S. Security Clearance
- Bachelor's Degree
- 9+ years of experience with DevOps, Site Reliability Engineering, software engineering, and/or cloud engineering
- Experience with GitLab, Azure DevOps and CI/CD
- Experience with technical leadership
- Experience with designing and implementing scalable computing infrastructure for data solutions, including cloud architectures (AWS, Azure, Google Cloud)