Senior Systems Engineer (Hybrid or Remote)
About The Role
Architecture & Delivery Lead design and implementation of platform solutions that meet organizational outcomes and non-functional requirements. Contribute to reference architectures and review designs for reliability, performance, and cost within the organization. Drive deep-dive investigations and root-cause analyses for complex issues; implement preventative patterns and guardrails. Collaborate with developers and clients to deliver solutions on Windows and Linux platforms, including .NET-based workloads. Manage and improve the engineering toolchain for the org (source control, CI/CD, packaging, artifact management).
Hybrid & Multi-Cloud
- Design and operate solutions across AWS, a second cloud (Azure or GCP), and on-prem environments for organizational workloads
- Engineer resilient hybrid connectivity and data services (e.g., transit gateways, private link, SD-WAN, SAN/NAS/NVMe-oF where applicable)
- Recommend workload placement based on performance, cost, compliance, and latency; document decision records for the org
- Implement consistent identity, secrets, and policy patterns across clouds and data centers
Automation & AI Enablement
- Implement Infrastructure as Code and configuration management to deliver repeatable, auditable environments
- Use AI copilots and agents to speed scripting, documentation, testing, incident triage, and change planning while validating outputs
- Build or integrate AI-powered internal tools (chatops, knowledge retrieval, automated runbooks) with appropriate logging and guardrails
- Promote safe AI use practices including data classification, prompt hygiene, redaction, and auditability within the organization
Reliability, Security & Operations
- Define and track SLIs/SLOs for owned platforms; build actionable alerts and dashboards for organizational visibility
- Perform capacity planning, performance engineering, resilience testing, and disaster recovery exercises for org services
- Apply security best practices with least privilege, encryption, patching baselines, and policy-as-code in collaboration with Security and Risk
- Provide tier-3 support, participate in on-call rotations, and continuously improve incident response with automation
Organization Contribution & Leadership
- Create reusable modules, images, and patterns that other teams in the organization can adopt
- Lead org-level working sessions, brown bags, and Communities of Practice to share knowledge and uplift standards
- Optimize cost and efficiency for org platforms (rightsizing, autoscaling, reservation strategies, lifecycle policies) and share outcomes
- Mentor systems engineers; provide technical coaching, code and design reviews, and support onboarding
About You
- 5+ years of experience in workstation or server administration
- 3+ years of experience in systems engineering delivering production solutions
- Bachelor's degree in computer science, information technology, or a related field (or equivalent experience)
- Proficiency with Microsoft Office; familiarity with documentation and work management tools (e.g., Confluence/Jira)
- Strong scripting/automation skills (PowerShell, Bash) and at least one higher-level language (.NET, Python, or Go)
- Solid knowledge of Windows and Linux servers, virtualization, and containerization (e.g., Kubernetes/Docker)
- Knowledge of networks including SAN/LAN, load balancing, and hybrid connectivity (VPN/Direct Connect/ExpressRoute)
- Experience operating in AWS and at least one additional cloud (Azure or GCP) plus on-prem data center environments
- Practical experience using AI-powered tools (copilots, LLM automations, chatops) with attention to security and data governance
- Infrastructure as Code skills (Terraform, CloudFormation/Bicep) and CI/CD for infrastructure and configurations
- Observability tooling experience (logs, metrics, traces) and reliability concepts (SLOs, error budgets)
- Knowledgeable in code development practices or equivalent enterprise application integration experience
- Experience supporting production infrastructure in hybrid environments, including cloud and on-premises systems, with a working understanding of networking, identity, access, and connectivity fundamentals
- Demonstrated automation mindset, including the ability to identify manual or repetitive processes and improve them through scripting, Infrastructure as Code, configuration management, CI/CD tooling, or other repeatable solutions
- Ability to apply security, access control, change management, documentation, validation, and operational readiness practices when delivering infrastructure changes
- Experience supporting data platforms, analytics infrastructure, ETL/ELT systems, reporting platforms, data lakes, data warehouses, or high-throughput database environments
- Experience building reusable infrastructure patterns, self-service workflows, automated runbooks, service catalog items, observability capabilities, Kubernetes/container platforms, or other internal platform capabilities
- Experience using AI-assisted engineering tools to support scripting, documentation, troubleshooting, testing, incident triage, or change planning while protecting sensitive data
Pay
$84,000.00-$207,000.00. The position may also be eligible for an annual bonus, incentives, and other employment-related benefits.
Schedule
This role may include participation in an on-call rotation to support production systems and ensure service reliability. On-call responsibilities may include coverage during nights and weekends. Frequency and scheduling will be determined by team needs.
This response is AI-generated, for reference only.