Sr. Software Engineer – Cloud Infrastructure and Devops
PayPal has been revolutionizing commerce globally for more than 25 years, creating innovative experiences that make moving money, selling, and shopping simple, personalized, and secure. We empower consumers and businesses in approximately 200 markets to join and thrive in the global economy. Our global, two-sided network connects hundreds of millions of merchants and consumers, enabling transactions and payments online or in person. We provide proprietary payment solutions, offer flexible funding sources, and enable safer, simpler fund transfers through products like PayPal, Venmo, and Xoom. Our core values—Inclusion, Innovation, Collaboration, and Wellness—guide our work as one global team with customers at the center.
About the role
This role delivers complete solutions spanning all phases of the Software Development Lifecycle (SDLC). You will be a key contributor to Venmo’s chaos engineering and business continuity efforts, building and operating systems that ensure our infrastructure can withstand and recover from any failure scenario. The Business Continuity team designs, builds, and operates infrastructure, tooling, and processes that protect Venmo from disruption, including disaster recovery automation, chaos engineering, incident game days, and backup and restore systems.
Responsibilities
- Deliver complete solutions spanning all phases of the SDLC (design, implementation, testing, delivery, and operations).
- Advise management on project-level issues and guide junior engineers.
- Operate with little day-to-day supervision, making technical decisions based on internal conventions and industry best practices.
- Act as a hands-on contributor while leading by example and mentoring junior engineers.
- Drive the scalability and reliability of Venmo’s AWS cloud infrastructure.
- Define, design, and implement solutions, coordinating stakeholders, navigating risks, and ensuring timely, high-quality delivery.
- Troubleshoot incidents, identify root causes, fix and document problems, and implement preventive measures.
- Develop and improve tools and automation to manage infrastructure and application configuration as code.
- Enhance the quality, reliability, and stability of infrastructure and operations.
- Design, implement, and operate chaos engineering experiments to proactively identify and remediate system weaknesses.
- Build and maintain disaster recovery automation, runbooks, and infrastructure across Venmo’s cloud environment.
- Plan and execute chaos and incident game day exercises to validate system resilience and team preparedness.
- Develop and maintain backup and restore tooling and processes for critical systems and data.
- Define and track resilience metrics, SLOs, and recovery time objectives for owned systems.
Requirements
- Bachelor’s degree in computer science or a related field.
- 5+ years of experience in software development or a related field.
- 3+ years of experience operating distributed applications 24x7x365 as part of a Cloud Engineering, DevOps, and/or SRE team.
- Extensive hands-on experience designing, implementing, and supporting infrastructure (AWS experience preferred) for global-scale services.
- Deep hands-on experience with IaaS and PaaS solutions from AWS (or similar cloud provider).
- Hands-on programming and scripting experience (Python, Java, Bash, Go).
- Hands-on experience with containers and container orchestration: Docker, Kubernetes.
- Strong communication skills with the ability to explain technical issues to a non-technical audience.
Preferred Qualifications
- Experience with chaos engineering tools and frameworks (e.g., AWS Fault Injection Simulator, Gremlin, Chaos Monkey, Litmus).
- Hands-on experience designing and implementing disaster recovery solutions for distributed systems.
- Experience developing and maintaining backup and restore tooling and strategies.
- Experience planning and executing game day exercises or incident simulations.
- Understanding of RTO/RPO requirements and how to design systems to meet them.
Our technical stack: AWS, EKS, Docker, GitHub Enterprise, Terraform, GitHub Actions, DataDog, Bash, Python, Go.
Pay
The base pay for this role depends on location and relevant experience. The expected pay ranges are:
- Austin, Texas: $130,500.00 - $193,600.00 annually
- San Jose, California: $143,500.00 - $212,850.00 annually
- Chicago, Illinois: $130,500.00 - $193,600.00 annually
- Scottsdale, Arizona: $123,500.00 - $183,700.00 annually
Additional compensation may include an annual performance bonus, equity, or other incentive compensation, as applicable.
Schedule
For the majority of employees, PayPal’s balanced hybrid work model offers 3 days in the office for in-person collaboration and 2 days at your choice of either the PayPal office or your home workspace.
Benefits
PayPal offers comprehensive, choice-based programs to support physical, emotional, and financial wellbeing, including:
- Generous paid time off.
- Healthcare coverage for you and your family.
- Resources to create financial security and support mental health.