Senior Platform Engineer
Leap · United States · 4 wk ago
RemoteRemoteEngineeringFull-time
Responsibilities
- Design, build, and operate secure, scalable, multi-account AWS infrastructure using Terraform, with reusable modules that drive consistency and reduce configuration drift.
- Evolve our CI/CD pipelines and deployment automation to enable faster, safer, and more reliable releases.
- Build self-service platforms and automation that reduce developer friction and manual operational toil.
- Implement monitoring, observability, and alerting that improve operational visibility and incident readiness.
- Drive AWS cost optimization and operational efficiencies without slowing engineering velocity.
- Partner closely with engineering teams to understand their needs, remove friction from their workflows, and turn recurring pain points into self-service solutions.
Requirements
- Minimum of Associate’s Degree, Bachelor’s Degree preferred.
- 7+ years in platform, infrastructure, DevOps, or SRE, with deep hands-on AWS.
- Strong expertise with Terraform and multi-account, multi-environment IaC patterns at scale, including remote state management (S3/DynamoDB).
- Strong experience with modern CI/CD systems and automated deployment pipelines (GitHub Actions preferred).
- Experience with AWS compute: EC2 and containerized workloads on ECS/Fargate (Docker).
- Experience supporting production databases such as MongoDB, PostgreSQL, and MySQL.
- Experience with observability platforms: CloudWatch and New Relic or similar.
- Strong software engineering skills in Python or TypeScript: you build platform tooling and automation as production-grade software (tested, versioned, operable), not one-off scripts.
- You think in distributed systems: failure modes, idempotency, retries, and backpressure, and how infrastructure design decisions shape application reliability at scale.
- Experience supporting SOC 2 audits and compliance in regulated environments. The bar is higher than using AI tools day to day. Specifically: You delegate whole pieces of work to autonomous agents, run several in parallel, and review and direct their output rather than hand-typing most of it. Infrastructure code is no exception.
- You know how to operationalize AI-generated changes in production environments: policy checks, plan-review gates, and blast-radius controls, so agent speed never outruns safety.
- Experience with AI-enabled operational workflows and LLM platforms such as Amazon Bedrock.
Preferred Skills
- Internal developer platform experience, including self-service infrastructure for engineering teams.
- EKS / Kubernetes experience: a direction we’re beginning to explore.
- A track record of modernizing manual, older deployment and infrastructure patterns toward fully automated models.
- Experience thriving in a high-autonomy, remote-first engineering culture.
- Experience with secure SDLC and software supply-chain security (e.g., Dependabot, push protection, dependency and secret scanning).
- Familiarity with AWS security posture tooling (GuardDuty, Security Hub, Config) is a plus.