Sr. Site Reliability Engineer (US Federal)
About the Role
This role will support one or more direct or indirect contracts with the U.S. Federal Government which, due to federal government security requirements, mandates that all Workday personnel working on the contracts be United States citizens (naturalized or native). You will be a key contributor on the Analytics Delivery Engineering team. As an SRE, you will help in building a clean, scalable, reliable, and automated services framework.
What You Will Do: Work on projects ranging from building out robust CI/CD to automation for both the Prism engineering team and SRE team, work on enhancements and improvements for monitoring, alerting, and tracing of not only internal services but most importantly, production services, collaborate with others on the SRE team to help set technical direction while ensuring requirements are met, help enhance, build, secure and maintain environments for Prism Analytics in Workday Federal deployments, interact and be the primary contact with multiple teams both internally in Workday Prism Analytics, participate in and facilitate production on-call duties and events to ensure reliability of the analytics applications and play a pivotal role in building the foundation that delivers Workday Analytics to GovCloud.
Responsibilities
- Design and implement: Design, develop, and implement solutions for our Kubernetes platform, including infrastructure automation, CI/CD pipelines, and observability tools.
- Build and maintain: Build and maintain core platform components, ensuring high availability, scalability, and security.
- Automate and optimize: Automate infrastructure provisioning, configuration management, and application deployments using tools like Terraform and Argo CD.
- Troubleshooting and support: Provide support and troubleshooting for platform-related issues, working closely with development teams to resolve problems.
- Security and compliance: Implement and maintain security best practices for the platform, ensuring compliance with industry standards.
- Documentation and knowledge sharing: Create and maintain comprehensive documentation for platform components and processes. Actively participate in knowledge sharing within the team.
- Collaboration: Collaborate effectively with other engineers, development teams, and stakeholders across multiple locations and time zones.
- Stay current: Stay up to date with the latest technologies and trends in the platform engineering space.
Qualifications
Basic Qualifications: 5+ years of hands-on experience working with infrastructure, either on-premises or cloud-based, with a deep understanding of systems architecture, networking, and security. Strong understanding of Kubernetes concepts, architecture, and administration. Bachelor's degree in a computer related field or equivalent work experience.
- Infrastructure as code: Proficiency in infrastructure automation tools like Terraform.
- CI/CD: Experience with building or maintaining CI/CD pipelines and tools like Argo CD.
- Cloud experience: 3+ years experience working with AWS cloud services in a production setting.
- Programming skills: Proficiency in at least one programming language, preferably GoLang or Python.
- Problem-solving: Strong analytical and problem-solving skills.
- Communication: Excellent communication and collaboration skills.
Nice to haves: Experience with monitoring and observability tools (Prometheus, Grafana), experience with security auditing and compliance frameworks, experience with Apache Spark, Docker, Kubernetes.
Pay
The annualized base salary ranges for the primary location and any additional locations are listed below. Workday pay ranges vary based on work location. As a part of the total compensation package, this role may be eligible for the Workday Bonus Plan or a role-specific commission/bonus, as well as annual refresh stock grants. Primary Location: USA.VA.Reston Primary Location Base Pay Range: $151,500 USD - $227,300 USD Additional US Location(s) Base Pay Range: $137,100 USD - $243,600 USD.
Schedule
Our Approach to Flexible Work With Flex Work, we’re combining the best of both worlds: in-person time and remote. Our approach enables our teams to deepen connections, maintain a strong community, and do their best work. We know that flexibility can take shape in many ways, so rather than a number of required days in-office each week, we simply spend at least half (50%) of our time each quarter in the office or in the field with our customers, prospects, and partners (depending on role).
Benefits
For more information regarding Workday’s comprehensive benefits, please click here.