Sr Manager, Platform DevOps
About the role
This role leads globally distributed DevOps/SRE teams across the US and India, with end-to-end accountability for workforce planning, team performance, and the hiring, development, and retention of a high-performing organization. It oversees the reliability, scalability, and cost efficiency of production and non-production environments across AWS and Azure, applying expertise in capacity planning, traffic management, and cloud optimization. Leading teams of 20+ engineers and contractors, the role drives platform delivery, technical and security enhancements, and multi-functional collaboration. Success is measured by platform reliability, timely delivery of capabilities, team growth, and the overall impact on organizational performance and customer experience.
Responsibilities
- Lead and manage distributed DevOps/SRE teams (US and India) globally, ensuring effective workforce planning, shift and availability management, performance development, mentorship, and continuous skill growth aligned with organizational needs.
- Own the security and vulnerability management lifecycle, ensuring timely remediation, cloud posture hardening, secure configuration management, and alignment with enterprise security, governance, and risk controls.
- Lead implementation of observability platforms across monitoring, logging, tracing, and alerting; develop dashboards and insights to proactively identify failures, bottlenecks, and performance deviations.
- Define and implement continuous improvement practices across technical fields and organizational processes. Drive SRE frameworks, including SLA/SLI/SLO definitions, reliability measurement, error-budget policies, and adoption of standards that improve operational excellence.
- Provide end-to-end ownership of incident management, including response coordination, root-cause analysis (RCA), post-incident reviews, and implementation of corrective actions to strengthen system resilience.
- Oversee technical vendor relationships to incorporate feature and function requests into product releases.
- Drive and maintain the current and future technical roadmap in collaboration with design and architecture teams.
- Collaborate with product, architecture, quality, and security organizations to align technical priorities and delivery objectives; drive execution of a long-term platform engineering roadmap covering modernization, automation, migrations, and innovation initiatives.
- Recruit and hire qualified managers and team members to strengthen the platforms and the support model.
Requirements
- Bachelor's Degree plus 7 years of related work experience OR a combination of education and experience deemed equivalent. Acceptable areas of study include Computer Science, Engineering, IT or equivalent experience. (Required)
- 7-10 years relevant Product Management experience in an agile software product development environment. (Required)
- 2-4 years experience in a leadership role. (Required)
- 7-10 years Technical Leadership: Strong command of cloud infrastructure (AWS & Azure), CI/CD systems, GitLab administration, IaC tools (Terraform/CloudFormation/Bicep), automation, and modern DevOps/SRE methodologies. (Preferred)
- 2-4 years experience managing teams of 5 or more resources in direct reporting relationships in a Platform Management organization. (Preferred)
Skills
- Strong understanding of Software Development Life Cycle (SDLC) and Agile methodologies
- Experience delivering complex technology initiatives across engineering and operations
- Expertise in vulnerability management, cloud security procedures, secure SDLC, compliance frameworks, and regulatory alignment
- Knowledge of observability concepts including monitoring, logging, and alerting
- Understanding of SLAs, SLOs, and service performance management
- Ability to collaborate with multi-functional partners and influence technical decisions
- Strong written and verbal communication skills with the ability to convey technical concepts clearly
- Analytical skills to assess system performance, operational metrics, and improvement opportunities
Qualifications
- Cloud certifications (AWS or Azure)
- Kubernetes or related containerization certifications
- At least 18 years of age
- Legally authorized to work in the United States
Travel may be required.
Pay
Base pay range: $160,000 - $288,500. The successful candidate’s actual pay will be based on work location, qualifications, and experience. Most Corporate employees are eligible for a year-end bonus based on company and/or individual performance, set at a percentage of the employee’s eligible earnings in the prior year (e.g., 20% target).
Benefits
- Medical, dental, and vision insurance
- Flexible spending account
- 401(k) with company match
- Annual stock grant and employee stock purchase plan
- Paid time off (about 4 weeks for new full-time employees, ~2.5 weeks for part-time) and up to 12 paid holidays annually
- Paid parental and family leave
- Family building benefits
- Backup care and enhanced family support
- Childcare subsidy
- Tuition assistance and college coaching
- Short- and long-term disability coverage
- Voluntary AD&D, accident, life, and long-term care insurance
- Mobile service and home internet discounts
- Pet insurance
- Commuter and transit programs
- Free, year-round money coaches