Direct Indexing Cloud Platform Infrastructure & Dev Ops Engineer
Vanguard · Oakland, CA · Yesterday
HybridEngineeringFull-time
Core Responsibilities
- Design, build, and operate cloud-native infrastructure supporting VPI's investment platform.
- Lead AWS platform architecture, modernization, and migration initiatives.
- Design scalable, highly available, and secure infrastructure patterns across development, test, and production environments.
- Manage Kubernetes and container platforms supporting distributed workloads.
- Drive infrastructure standardization, resiliency, disaster recovery, and capacity planning.
- Continuously improve platform observability, monitoring, logging, and operational readiness.
DevOps & Automation
- Establish and evolve CI/CD pipelines and engineering automation practices.
- Drive Infrastructure-as-Code adoption using modern provisioning and configuration-management frameworks.
- Automate deployment, environment management, compliance validation, and operational processes.
- Improve engineering productivity through tooling, platform services, and self-service capabilities.
- Define and implement operational best practices for cloud-native application delivery.
Security Engineering & Architecture Enablement
- Partner with Enterprise Security and application teams to design secure platform architectures.
- Enable implementation of security controls, vulnerability remediation, identity management, encryption, network segmentation, and cloud governance frameworks.
- Support regulatory, audit, risk, and compliance initiatives.
- Drive proactive remediation of security findings and operational risks.
- Establish secure-by-default platform patterns and guardrails.
Enterprise & External System Integration
- Design and support integrations with Vanguard enterprise platforms and external vendors.
- Build secure and scalable connectivity patterns across cloud and on-premises environments.
- Enable data exchange, APIs, messaging, file transfer, and service integrations.
- Partner with business and technology stakeholders to evaluate emerging technology solutions and integration opportunities.
Innovation, Architecture & Proof of Concepts
- Lead technical evaluations, proofs of concept, and pilot initiatives supporting strategic platform evolution.
- Assess new technologies, cloud services, engineering tools, and architecture patterns.
- Collaborate with product, research, and engineering teams to accelerate innovation.
- Influence long-term platform architecture and modernization roadmaps.
Operational Excellence
- Serve as a senior escalation point for complex production issues and platform incidents.
- Lead root cause analysis, remediation planning, and reliability improvements.
- Establish service-level objectives, operational metrics, and reliability standards.
- Reduce operational risk through automation, standardization, and improved engineering practices.
- Ensure platform services meet performance, availability, scalability, and recoverability objectives.
Technical Leadership
- Provide architecture guidance and technical leadership across infrastructure and platform domains.
- Mentor engineers and promote engineering excellence throughout the organization.
- Lead design reviews and infrastructure governance processes.
- Champion modern engineering practices and continuous improvement.
Qualifications
- 10+ years of experience in infrastructure engineering, platform engineering, DevOps, cloud engineering, or related fields.
- Deep expertise with AWS cloud services and cloud-native architectures.
- Strong experience operating Kubernetes-based platforms in production environments.
- Expertise in Infrastructure-as-Code and automation frameworks.
- Experience designing secure, highly available distributed systems.
- Strong Linux, networking, and systems engineering background.
- Experience implementing CI/CD pipelines and modern DevOps practices.
- Experience supporting mission-critical production systems with high availability requirements.
- Strong understanding of information security, cloud security, and risk management principles.
- Proven ability to lead complex technical initiatives across multiple teams.
- Excellent troubleshooting, problem-solving, and communication skills.
Preferred Experience
- Within asset management, wealth management, trading, or investment platforms.
- Experience with Terraform, Kubernetes (EKS), Docker, GitHub Actions, Jenkins, or similar technologies.
- Experience supporting enterprise cloud migrations and large-scale modernization programs.
- Experience with observability platforms, monitoring solutions, and SRE practices.
- Familiarity with financial services regulatory and compliance requirements.
- Experience integrating cloud platforms with enterprise systems and third-party vendors.