Sr. Site Reliability Engineer
$110,000 - $145,000 / year + Bonus
About the role
The insurance industry runs on Vertafore. We equip agencies, MGAs, and carriers with the core digital systems, specialized AI, and data-driven foundation to eliminate distribution drag across the insurance lifecycle, spanning sales, servicing, and back-office operations. Underpinned by unmatched speed and performance power, we are the trusted backbone that’s taking the insurance industry from friction to flow with Distribution Velocity—speed, performance, and trust—to drive growth at scale.
With over 95% of the top agencies and insurers and 50% of industry compliance transactions running through Vertafore, we lead at the intersection of innovation and trust, giving insurance professionals the confidence to transform and win in the AI era. Our reach is global, with headquarters in Denver, Colorado, and offices across the U.S., Canada, and India.
We are seeking a Senior Site Reliability Engineer to own the reliability, scalability, performance, and operational integrity of critical production services. This role is accountable for the full-service lifecycle, from design and deployment readiness through production operations, incident response, and continuous improvement. Reliability is a core engineering responsibility, requiring strong software engineering skills and autonomous operation across AWS, hybrid data centers, and customer-hosted environments.
Responsibilities
- Own production services end to end. Accountable for reliability, availability, scalability, performance, and operational health.
- Define and manage SLIs and SLOs, using error budgets to guide delivery decisions.
- Influence service and system design to improve fault tolerance, observability, and operational sustainability.
- Debug complex production issues across application code, services, and infrastructure using software engineering practices.
- Perform root cause analysis using logs, metrics, traces, and code-level investigation.
- Build automation and self-healing mechanisms to prevent repeat failures.
- Execute production changes (patching, certificate management, software releases) with safety, automation, and observability.
- Design and operate production observability aligned to service health and customer impact.
- Lead and participate in incident response for high-severity events.
- Collaborate with engineering, product, architecture, and operations teams.
- Operate with autonomy and sound judgment in reliability decisions.
- Participate in an on-call rotation; flexible hours as required.
Requirements
- 8+ years of hands-on Site Reliability Engineering or reliability-focused engineering experience with end-to-end service ownership.
- Proven operation at a senior engineering scope with accountability for reliability outcomes.
- Strong software engineering skills in C#, .NET, Java, Python, React, or similar technologies.
- Practical experience applying SRE principles (SLIs, SLOs, error budgets).
- Hands-on experience with AWS, Kubernetes, CI/CD, infrastructure as code, and hybrid environments.
- Strong knowledge of Linux and Windows systems, application platforms, and relational databases.
- Bachelor’s or master’s degree in computer science or equivalent experience.
Skills
- A fast learner and problem solver.
- Ability to document procedures.
- Able to meet deadlines.
- Good communication skills, able to deliver messages effectively to technical and non-technical audiences.
- Able to comply with processes and procedures.
- Able to maintain professional composure in any situation.
- Flexible in working extended hours on occasions or as required.
- Exposure to the insurance industry is desired but not mandatory.
- Driven to improve, personally and professionally.
- Operate best in a fast-paced, flexible work environment with the ability to work in a team.
Additional Requirements
- High-speed internet to accommodate working from home needs.
- Occasional travel to our office location is required.
- Occasional lifting and/or moving up to 10 pounds.
- Frequent repetitive hand and arm movements required to operate a computer.
- Specific vision abilities required by this job include close vision (working on a computer, etc.).
- Frequent sitting and/or standing.
Benefits
Canada Only
- Medical, vision & dental plans
- Life, AD&D
- Short Term and Long Term Disability
- Pension Plan & Employer Match
- Maternity, Paternity, and Parental Leave
- Employee and Family Assistance Program (EFAP)
- Education Assistance
- Additional programs - Employee Referral and Internal Recognition
US Only
- Flexible First work environment (hybrid/remote options)
- Medical, vision & dental plans (PPO & high-deductible options)
- Health Savings Account & Flexible Spending Accounts (Health Care FSA, Dental & Vision FSA, Dependent Care FSA, Commuter FSA)
- Life, AD&D (Basic & Supplemental), and Disability
- 401(k) Retirement Savings Plan & Employer Match
- Supplemental Plans - Pet insurance, Hospital Indemnity, and Accident Insurance
- Parental Leave & Adoption Assistance
- Employee Assistance Program (EAP)
- Education & Legal Assistance
- Tuition Reimbursement
- Employee Referral and Internal Recognition programs
- Wellness programs
- Commuter Benefits (Denver)
Bonus plans:
- The Professional Services (PS) and Customer Success (CX) bonus plans are a quarterly monetary bonus based on individual and practice performance against specific business metrics.
- The Vertafore Incentive Plan (VIP) is an annual monetary bonus for eligible employees based on both individual and company performance.
- Commission plans are tailored to each sales role but include components such as quota, MBOs, and ABPMs.