Sr Manager, Platform DevOps
About the role
This role leads globally distributed DevOps/SRE teams across the US and India, with end-to-end accountability for workforce planning, team performance, and the hiring, development, and retention of a high-performing organization. It oversees the reliability, scalability, and cost efficiency of production and non-production environments across AWS and Azure, applying expertise in capacity planning, traffic management, and cloud optimization. Leading teams of 20+ engineers and contractors, the role drives platform delivery, technical and security enhancements, and multi-functional collaboration. Success is measured by platform reliability, timely delivery of capabilities, team growth, and the overall impact on organizational performance and customer experience.
Responsibilities
- Lead and manage distributed DevOps/SRE teams (US and India) globally, ensuring effective workforce planning, shift and availability management, performance development, mentorship, and continuous skill growth aligned with organizational needs.
- Own the security and vulnerability management lifecycle, ensuring timely remediation, cloud posture hardening, secure configuration management, and alignment with enterprise security, governance, and risk controls.
- Lead implementation of observability platforms across monitoring, logging, tracing, and alerting; develop dashboards and insights to proactively identify failures, bottlenecks, and performance deviations.
- Define and implement continuous improvement practices across technical fields and organizational processes. Drive SRE frameworks, including SLA/SLI/SLO definitions, reliability measurement, error-budget policies, and adoption of standards that improve operational excellence.
- Provide end-to-end ownership of incident management, including response coordination, root-cause analysis (RCA), post-incident reviews, and implementation of corrective actions to strengthen system resilience.
- Oversee technical vendor relationships to incorporate feature and function requests into product releases.
- Drive and maintain the current and future technical roadmap in collaboration with design and architecture teams.
- Collaborate with product, architecture, quality, and security organizations to align technical priorities and delivery objectives; drive execution of a long-term platform engineering roadmap covering modernization, automation, migrations, and innovation initiatives.
- Recruit and hire qualified managers and team members to strengthen the platforms and the support model.
Requirements
- Bachelor's Degree plus 7 years of related work experience OR a combination of education and experience deemed equivalent. Acceptable areas of study include Computer Science, Engineering, IT or equivalent experience. (Required)
- 7-10 years relevant Product Management experience in an agile software product development environment. (Required)
- 2-4 years experience in a leadership role. (Required)
- 7-10 years technical leadership with strong command of cloud infrastructure (AWS & Azure), CI/CD systems, GitLab administration, IaC tools (Terraform/CloudFormation/Bicep), automation, and modern DevOps/SRE methodologies. (Preferred)
- 2-4 years experience managing teams of 5 or more resources in direct reporting relationships in a Platform Management organization. (Preferred)
Skills
- Strong understanding of Software Development Life Cycle (SDLC) and Agile methodologies.
- Experience delivering complex technology initiatives across engineering and operations.
- Expertise in vulnerability management, cloud security procedures, secure SDLC, compliance frameworks, and regulatory alignment.
- Knowledge of observability concepts including monitoring, logging, and alerting.
- Understanding of SLAs, SLOs, and service performance management.
- Ability to collaborate with multi-functional partners and influence technical decisions.
- Strong written and verbal communication skills with the ability to convey technical concepts clearly.
- Analytical skills to assess system performance, operational metrics, and improvement opportunities.
Qualifications
- Cloud certifications (AWS or Azure).
- Kubernetes or related containerization certifications.
- At least 18 years of age.
- Legally authorized to work in the United States.
Travel may be required.
Pay
Base pay range: $160,000 - $288,500. The successful candidate’s actual pay will be based on work location, qualifications, and experience. Corporate bonus target: 20%. Most corporate employees are eligible for a year-end bonus based on company and/or individual performance.
Benefits
- Medical, dental, and vision insurance.
- Flexible spending account.
- 401(k) with company match.
- Annual stock grant and employee stock purchase plan.
- Paid time off and up to 12 paid holidays (about 4 weeks for new full-time employees, 2.5 weeks for new part-time employees annually).
- Paid parental and family leave.
- Family building benefits, back-up care, enhanced family support, and childcare subsidy.
- Tuition assistance and college coaching.
- Short- and long-term disability.
- Voluntary AD&D, accident, life, disability, and long-term care insurance.
- Mobile service and home internet discounts.
- Pet insurance.
- Commuter and transit programs.
- Free, year-round money coaches.