Jobs · North Carolina

Engineering Manager, Observability Platforms

TEKsystems · Raleigh, NC · 2 wk ago
Hybrid$169k–$200k/yrFull-time

We are seeking an experienced Observability Manager to lead our enterprise observability strategy and capabilities. This individual will drive visibility, reliability, and operational excellence across a large-scale technology environment. The ideal candidate is a strong technical leader with experience building and managing observability programs, establishing monitoring standards, and partnering with engineering teams to improve system health and performance.

Responsibilities

  • Leadership & Strategy
    • Lead and mentor observability engineers and technical teams.
    • Define and execute the enterprise observability roadmap.
    • Establish standards, governance, and best practices for monitoring and telemetry.
    • Partner with engineering, operations, architecture, and leadership teams to improve operational visibility and reliability.
    • Drive adoption of observability tools and practices across the organization.
  • Observability & Reliability
    • Develop enterprise-wide monitoring strategies covering metrics, logs, traces, and alerting.
    • Champion modern observability practices and instrumentation standards.
    • Improve visibility across development, test, and production environments.
    • Support incident response, root cause analysis, and reliability improvement initiatives.
    • Define KPIs and success metrics that demonstrate platform effectiveness and business value.
  • Operational Excellence
    • Optimize observability platforms for performance, scalability, and cost efficiency.
    • Establish telemetry governance and data ingestion best practices.
    • Create dashboards, reporting, and operational insights for technical and business stakeholders.
    • Drive continuous improvement initiatives focused on system reliability and operational readiness.

Requirements

  • Experience leading observability, SRE, platform engineering, DevOps, or software engineering teams.
  • Strong understanding of:
    • Metrics
    • Logs
    • Distributed tracing
    • Monitoring and alerting strategies
    • Incident management and operational visibility
  • Experience with enterprise observability platforms such as:
    • Datadog
    • Splunk
    • Dynatrace
    • New Relic
    • Grafana
    • AppDynamics
    • Coralogix
    • Similar observability solutions
  • Experience with cloud environments, preferably AWS and/or Azure.
  • Strong stakeholder management and communication skills.
  • Ability to influence technical direction and drive adoption across multiple teams.

Qualifications

  • Preferred Qualifications
    • Experience with OpenTelemetry, Prometheus, or similar observability frameworks.
    • Experience developing observability standards and governance models.
    • Experience leading geographically distributed teams.
    • Background in reliability engineering, operational analytics, or platform operations.
    • Experience managing large-scale enterprise monitoring deployments.

What Success Looks Like

  • Improved operational visibility across the enterprise.
  • Increased monitoring coverage and observability adoption.
  • Reduced incident detection and resolution times.
  • Strong governance around telemetry collection and platform utilization.
  • Improved reliability, performance, and operational readiness across critical applications.
  • High-performing, engaged observability teams aligned with business objectives.

Why Join Us?

This is an opportunity to lead a highly visible function that directly impacts application reliability, customer experience, and operational excellence. You'll help shape the future of enterprise observability while partnering with engineering leaders to drive measurable improvements across the organization.

Pay

The pay range for this position is $169,000.00 - $200,000/yr.

Benefits

  • 401k match up to 5%
  • 18 days PTO with an option to buy an additional week
  • Flexible hours and remote work options

Schedule

This is a hybrid position based in Raleigh, NC. Permanent role with no micromanagement, targeting experienced professionals.

Similar jobs