Jobs · Information Technology · Missouri

Senior Site Reliability Systems Engineer

Apex Systems · St Louis, MO · Yesterday
Information TechnologyFull-time

Overview

This position offers the opportunity to work fully remote within the United States (excluding Alaska and Hawaii). Candidates must be able to work within U.S. Central Time core business hours. Periodic travel to company locations for meetings, team events, or business needs may be required a few times per year.

Position Summary

The Rental Operations Products team is responsible for the development and support of mission-critical rental ticketing and operational systems. This role is responsible for implementing enhancements, monitoring system health, conducting capacity planning, resolving production issues, and designing solutions that support an API-first architecture strategy. The team partners closely with multiple technology groups, including mobile applications, kiosks, web platforms, and customer experience solutions, to ensure consistent, reliable service delivery. This is an opportunity to contribute to the evolution of highly available systems and the implementation of new capabilities that support business growth and operational excellence.

The Senior Systems Engineer (Engineer II) is responsible for ensuring the availability, performance, efficiency, monitoring, capacity planning, change management, and incident response of highly available production systems. These systems consist of a hybrid environment spanning cloud platforms, web services, and legacy technologies. The ideal candidate will have experience with Site Reliability Engineering (SRE) principles and a strong understanding of both software engineering and infrastructure operations. This role serves as a bridge between development and operations teams by applying engineering practices to system administration and operational support.

As a Senior Systems Engineer, you will be expected to proactively monitor solution health, ensure systems meet established service-level objectives, and assist in resolving performance issues, capacity concerns, and outage events. You will also be responsible for supporting maintenance activities such as upgrades and patching, while developing a comprehensive understanding of the overall solution ecosystem to drive reliability and continuous improvement. This role requires a technical leader who can serve as a subject matter expert, represent the team on complex initiatives, evaluate technology effectiveness, and recommend enhancements that improve system reliability, scalability, and consistency.

Responsibilities

  • Contribute to strategic capacity planning initiatives.
  • Focus on production infrastructure support and operational improvement efforts.
  • Monitor key performance metrics and proactively address issues.
  • Serve as a subject matter expert in multiple technical areas.
  • Provide technical guidance and leadership for projects and initiatives.
  • Support large-scale and complex assignments.
  • Operate with a high degree of autonomy while collaborating across teams.
  • Define, develop, communicate, and implement standards, processes, and procedures.
  • Build and maintain strong relationships with stakeholders across the organization.
  • Partner with architects to improve solution performance, scalability, reliability, and quality.
  • Create and maintain technical documentation.
  • Mentor and support less experienced team members.
  • Explore emerging technologies and contribute innovative solutions.
  • Participate in troubleshooting, root cause analysis, and incident resolution activities.
  • Support system health during maintenance windows, upgrades, and deployments.

Required Qualifications

  • Authorized to work in the United States without current or future sponsorship requirements.
  • Must reside within the United States (excluding Alaska and Hawaii).
  • Ability to work within U.S. Central Time core business hours.
  • 3+ years of experience supporting software development environments utilizing technologies such as Java, Web Services, C/C++, and PL/SQL.
  • 2+ years of experience configuring and supporting middleware technologies including Tuxedo, WebLogic, and Tomcat.
  • 2+ years of experience administering AIX and/or Linux systems.
  • 2+ years of experience developing scripts for automation, system administration, or application support using tools such as Shell/Bash, Python, Perl, PowerShell, or similar technologies.
  • 1+ year of experience with monitoring and observability platforms such as Splunk, Dynatrace, or comparable tools.
  • Strong commitment to incorporating security best practices into daily responsibilities and technical decisions.

Preferred Qualifications

  • Bachelor's degree in Computer Science, Information Systems, Management Information Systems, or a related field.
  • General knowledge of network engineering concepts.
  • Experience with infrastructure and system testing methodologies.
  • Strong problem-solving and troubleshooting skills.
  • Familiarity with capacity planning techniques and methodologies.
  • Understanding of hardware and infrastructure concepts, including networking devices, resiliency patterns, and system topology.
  • Working knowledge of SQL.

Similar jobs

Senior Systems Engineer

Intuitive Research and Technology CorporationColorado Springs, CO· 2 mo ago
Information Technology$150k–$210k/yrapply on intuitive.wd1.myworkdayjobs.com

Senior Systems Engineer

Radiance TechnologiesHuntsville, AL· 3 mo ago
Information Technologyapply on radiancetech.wd12.myworkdayjobs.com