Senior Manager, Software Development Engineering
About the Role
The Manager, SRE & Cloud Engineering manages Cloud Engineering and Site Reliability Engineering teams responsible for cloud infrastructure, service reliability, automation, observability, and operational excellence to support the delivery of highly available business-critical applications.
Responsibilities
- Lead and develop engineering staff, setting clear performance expectations, supporting professional growth, and fostering a high-performance, accountable culture.
- Oversee day-to-day operations and technical escalations for Cloud Engineering & Architecture and Site Reliability Engineering staff and processes.
- Collaborate with Incident Management teams and leadership to build observability dashboards monitoring environmental and application health.
- Execute on cloud engineering and architecture direction by designing and managing scalable infrastructure across public cloud, driving IaC adoption, CI/CD (Azure/AWS), and automation best practices.
- Build and mature SRE practices, including defining and enforcing SLOs/SLIs, establishing observability (metrics, logs, tracing), automating incident response and runbooks, and reducing operational toil through tooling and process automation.
- Drive transparency into environment variables, enabling a metrics-centric operating model with executive visibility into system reliability, incident trends, infrastructure health, and team performance.
- Partner with information security to ensure cloud environments meet security, risk, and compliance requirements, implement guardrails, enforce policies, and maintain audit readiness.
- Hire, mentor, and develop a team of Site Reliability and Cloud Engineers, providing regular feedback and career growth guidance.
- Define and track key reliability metrics (SLIs, SLOs, error budgets) across services owned by the team.
- Partner with engineering leadership to influence system design decisions for scalability, resilience, and observability.
- Manage day-to-day workload and priorities for a team of SRE and cloud engineers.
Requirements
Required Education: Bachelor's degree in Computer Science, Information Technology, Engineering, or a related field (or 4 additional years of relevant experience in lieu of a degree).
Required Experience: 5+ years of progressive experience in technology operations, infrastructure engineering, cloud architecture, or site reliability engineering, with 3+ years in a people leadership role managing managers.
Skills
- Proven track record of leading and scaling high-performing engineering teams across distributed locations.
- Strong focus on accountability and transparency.
- Deep expertise in cloud platforms and modern infrastructure practices including IaC, containerization, and CI/CD pipelines.
- Strong understanding of SRE principles, observability tooling, and incident management platforms.
- Excellent communication skills, able to translate complex technical topics into clear, business-aligned strategies.
- Track record of driving operational efficiency through automation, process improvement, and metrics-driven decision making.
- Cloud architecture or ITIL certifications a plus.
Location
Marlboro or Chelmsford, MA
Pay
Target Compensation: $146,500 - $176,000 per year
Schedule
Monday - Friday, 8am - 5pm
Benefits
- Traditional medical, dental, and vision coverage.
- Generous 401(k) match.
- Paid Time Off: Up to 15 days accrued in your first year, plus 40 hours of sick time and 3 personal days (refreshed annually).
- Paid federal holidays.
- Special employee pricing on lending products such as mortgage, auto, and personal loans (eligibility subject to standard account requirements and underwriting criteria).