Site Reliability Engineer
Cisco · Appleton, WI · Today
Information Technology$165k–$241k/yrFull-time
What You’ll Do
- Use best practices and knowledge of internal or external business issues to improve products or services; works independently with minimal guidance and direction.
- Act as a resource for colleagues with less experience; understand project and/or department needs and establish relationships with appropriate cross-functional stakeholders to gather input, collect information, and complete work steps.
- Design and deploy small to mid-size or moderately complex solutions to optimize reliability, availability, latency, and performance.
- Integrate knowledge of design, automation, and deployment with expertise in coding to improve service reliability for existing or new systems and adapt for regions, countries, or customers.
- Design and test high availability and disaster recovery measures for our services to ensure automation is improving reliability, scalability, and velocity.
- Forecast and build reports to determine at what point resources will be at capacity.
- Design and implement tools that provide visibility into performance and reliability of our infrastructure; build automated platforms.
- Monitor the environment and work with Developers and Ops to identify problems and develop monitoring tools that provide visibility into performance and reliability.
- Serve as on-call SRE, lead post mortems, and write root cause analysis.
- Build and ensure security controls are in place in regards to architectural design; collaborate with security in designing or providing input to security controls; actively contribute in security incident response.
- Evaluate the scalability, resiliency, performance, and security properties and techniques used in production environments.
- Support uptime of production services through an On-Call rotation, including monitoring and alerting to meet internal Service Level Objectives (SLOs) and customer-facing Service Level Agreements (SLAs).
- Ensure reliable incident processes by conducting Disaster Recovery drills.
- Improve reliability through incident management by investigating incidents, implementing remediation strategies, and learning from past incidents to make improvements.
- Determine the reliability and security requirements of components and systems to meet the reliability objectives of the company, customers, and relevant governmental agencies.
- Reduce operational expenses through automation by identifying and mitigating failure points and automating repetitive and resource-intensive tasks.
- Develop new acceleration techniques and analytical tools to ensure early identification of potential issues with new products, packaging, processes, and overall product reliability.
- Ensure a reliable and scalable network; manage network/cloud infrastructure and storage systems supporting business operations; respond to planned maintenance, real-time outages, and issues.
- Plan, design, and implement local and wide-area network solutions between multiple platforms and protocols.
Minimum Qualifications
- Bachelors + 7 years of related experience, or Masters + 4 years of related experience, or PhD + 1 year of related experience.
- Solid conceptual and practical knowledge in primary technical job family and knowledge of related technical job families.
- Experience working with a range of technologies.
Pay & Benefits
Salary Range: $165,000.00 to $241,400.00 (projected range for new hires in U.S. and/or Canada locations, not including incentive compensation, equity, or benefits). Individual pay determined by hiring location, market conditions, job-related skillset, experience, qualifications, education, certifications, and/or training.
Full salary ranges by specific state:
- New York City Metro Area: $165,000.00 – $277,600.00
- Non-Metro New York State & Washington State: $146,700.00 – $247,000.00
Benefits (U.S. employees, subject to plan eligibility rules):
- Medical, dental, and vision insurance
- 401(k) plan with Cisco matching contribution
- Paid parental leave
- Short and long-term disability coverage
- Basic life insurance
- Eligibility to receive grants of Cisco restricted stock units (vest following continued employment for defined periods)
Paid time away (subject to Cisco policies):
- 10 paid holidays per full calendar year, plus 1 floating holiday for non-exempt employees
- 1 paid day off for employee's birthday
- Paid year-end holiday shutdown
- 4 paid days off for personal wellness
- Non-exempt employees: 16 days of paid vacation time per full calendar year (accrued at 4.92 hours per pay period for full-time)
- Exempt employees: Flexible vacation time off program (no defined limit, subject to availability and business limitations)
- 80 hours of sick time off provided on hire date and each January 1st thereafter; up to 80 hours of unused sick time carried forward year to year
- Additional paid time away may be requested for critical or emergency family issues
- Optional 10 paid days per full calendar year to volunteer
Additional compensation:
- Non-sales roles: Eligible to earn annual bonuses subject to Cisco's policies
- Sales roles: Performance-based incentive pay on top of base salary, split between quota and non-quota components. For quota-based incentive pay: 0.75% of incentive target for each 1% of revenue attainment up to 50% of quota; 1.5% of incentive target for each 1% of attainment between 50% and 75%; 1% of incentive target for each 1% of attainment between 75% and 100%; once performance exceeds 100% attainment, incentive rates are at or above 1% for each 1% of attainment with no cap. For non-quota-based sales performance elements (e.g., strategic sales objectives), Cisco may pay 0% up to 125% of target. Cisco sales plans do not have a minimum threshold of performance for sales incentive compensation to be paid.