Jobs · Engineering · Massachusetts

Senior Manager, Software Development Engineering

Digital Federal Credit Union · Chelmsford, MA · 1 wk ago
On-siteEngineering$147k–$176k/yrFull-time

About the Role

The Manager, SRE & Cloud Engineering manages Cloud Engineering and Site Reliability Engineering teams responsible for cloud infrastructure, service reliability, automation, observability, and operational excellence to support the delivery of highly available business-critical applications.

Responsibilities

  • Lead and develop engineering staff, setting clear performance expectations, supporting professional growth, and fostering a high-performance, accountable culture.
  • Oversee day-to-day operations and technical escalations for Cloud Engineering & Architecture and Site Reliability Engineering staff and processes.
  • Collaborate with Incident Management teams and leadership to build observability dashboards monitoring environmental and application health.
  • Execute on cloud engineering and architecture direction by designing and managing scalable infrastructure across public cloud, driving IaC adoption, CI/CD (Azure/AWS), and automation best practices.
  • Build and mature SRE practices, including defining and enforcing SLOs/SLIs, establishing observability (metrics, logs, tracing), automating incident response and runbooks, and reducing operational toil through tooling and process automation.
  • Drive transparency into environment variables, enabling a metrics-centric operating model with executive visibility into system reliability, incident trends, infrastructure health, and team performance.
  • Partner with information security to ensure cloud environments meet security, risk, and compliance requirements, implement guardrails, enforce policies, and maintain audit readiness.
  • Hire, mentor, and develop a team of Site Reliability and Cloud Engineers, providing regular feedback and career growth guidance.
  • Define and track key reliability metrics (SLIs, SLOs, error budgets) across services owned by the team.
  • Partner with engineering leadership to influence system design decisions for scalability, resilience, and observability.
  • Manage day-to-day workload and priorities for a team of SRE and cloud engineers.

Requirements

Required Education: Bachelor's degree in Computer Science, Information Technology, Engineering, or a related field (or 4 additional years of relevant experience in lieu of a degree).

Required Experience: 5+ years of progressive experience in technology operations, infrastructure engineering, cloud architecture, or site reliability engineering, with 3+ years in a people leadership role managing managers.

Skills

  • Proven track record of leading and scaling high-performing engineering teams across distributed locations.
  • Strong focus on accountability and transparency.
  • Deep expertise in cloud platforms and modern infrastructure practices including IaC, containerization, and CI/CD pipelines.
  • Strong understanding of SRE principles, observability tooling, and incident management platforms.
  • Excellent communication skills, able to translate complex technical topics into clear, business-aligned strategies.
  • Track record of driving operational efficiency through automation, process improvement, and metrics-driven decision making.
  • Cloud architecture or ITIL certifications a plus.

Location

Marlboro or Chelmsford, MA

Pay

Target Compensation: $146,500 - $176,000 per year

Schedule

Monday - Friday, 8am - 5pm

Benefits

  • Traditional medical, dental, and vision coverage.
  • Generous 401(k) match.
  • Paid Time Off: Up to 15 days accrued in your first year, plus 40 hours of sick time and 3 personal days (refreshed annually).
  • Paid federal holidays.
  • Special employee pricing on lending products such as mortgage, auto, and personal loans (eligibility subject to standard account requirements and underwriting criteria).

Similar jobs