Jobs · Engineering · South Carolina

COBOL, Mainframe & Cloud Manager Software Engineer- US

DXC Technology · Charleston, SC · 3 wk ago
EngineeringTemporary

About The Role

The Software Engineer is a senior, hands-on technical leader responsible for improving the reliability, availability, scalability, and performance of complex enterprise systems across both legacy and modern cloud platforms. This role emphasizes execution, deep technical expertise, and independent operation in high-stakes environments.

What You'll Do

  • Provides leadership in improving reliability and performance for targeted services, platforms, or programs
  • Operates across hybrid environments (including mainframe workloads and AWS-hosted services) with diverse integrations and dependencies
  • Drives incident reduction through measurable reliability goals (SLIs/SLOs), runbooks, and automation within defined engagement scopes
  • Influences technical standards, operational practices, and architecture within local domains, partnering with teams to implement durable reliability improvements

Who You Are

  • 7+ years’ hands on experience as Software Engineer – Performance and Reliability Engineer
  • Bachelors degree in an IT related field or minimum 7+ years of experience
  • Mainframe and COBOL application experience
  • Modern enterprise systems (e.g., Java-based, cloud-native, and other contemporary application platforms) and their operational characteristics
  • AWS environments, including QA, UAT, and Production, with strong understanding of networking, compute, and storage fundamentals
  • Performance engineering: profiling, load/stress testing, latency analysis, capacity planning, and tuning across application and infrastructure layers
  • Ability to guide performance profiling with application teams (e.g., CPU/memory profiling, thread/heap analysis, database/query tuning), identify problem code paths, and recommend targeted fixes, translating findings into prioritized remediation work
  • Regarded as a technical authority among peer engineers; strong communicator during incidents and while driving cross-team remediation
  • Calm, methodical execution in high-pressure situations; leads incident response, escalation, and post-incident reviews (blameless postmortems)
  • Improves operational readiness through runbooks and on-call execution, including responding whenever there is a fire
  • Expert in troubleshooting with observability data—metrics, logs, and traces—to isolate failure modes, quantify impact, and validate fixes
  • Ability to use observability and monitoring tools (such as Dynatrace) to instrument services, analyze end-to-end transactions, and pinpoint bottlenecks or failures in urgent and non-urgent scenarios
  • Skilled in diagnosing integration and dependency issues across enterprise platforms, identifying root causes in interconnected systems, and recommending durable fixes to restore and harden functionality
  • Experienced with hosting and infrastructure components (virtual machines, container orchestration, CI/CD pipelines, and network appliances), enabling comprehensive troubleshooting, resilience improvements, and automation
  • Frequently sought out for critical reliability or performance issues requiring immediate attention and clear leadership
  • Improves outcomes by reducing mean time to restore (MTTR) and recurring incidents through automation, runbooks, and durable fixes
  • Renowned for delivering measurable reliability and performance improvements (e.g., SLO attainment, latency reduction), not just frameworks or process artifacts

Work Environment

If you live within 40 km (25 miles) of a DXC office, you are expected to work onsite at least two days per week.

Similar jobs