Sr. Software Engineer, Observability - Slack
Salesforce is the #1 AI CRM, where humans and agents collaborate to drive customer success. The Monitoring Infrastructure team in the Service Delivery Platform & Reliability group at Slack develops platforms and tools that produce telemetry, provide insights, and improve observability in Slack’s production services, focusing on performance and reliability.
About The Team
The Monitoring Infrastructure team builds and maintains distributed services that process millions of data points per second, with self-healing and scaling capabilities. We work with open-source observability technologies like Astra and Prometheus, cloud providers such as AWS, and develop software using Go, Python, or Java. As part of this team, you’ll focus on log pipelines and collaborate with engineering, product development, and customer experience teams to provide insights that drive decisions and ensure a positive user experience for Slack customers.
Responsibilities
- Build, maintain, and ensure timely delivery of high-volume event log pipelines.
- Create libraries, tools, and automation to ensure critical event data reaches the right place.
- Encourage a culture of Observability at Slack by identifying problem areas and consulting on improving system visibility.
- Prototype tooling interfaces or build new features for engineering use cases.
- Improve auto-remediation in logging infrastructure to avoid recurring failures.
- Teach engineers or customer experience agents how to use tools to introspect their systems.
- Collaborate with the rest of Platform to integrate observability solutions, enabling data-driven decisions for operations, cost optimization, and scaling.
- Participate in the Monitoring Infrastructure on-call rotation, triaging, and addressing production issues as they arise.
Requirements
- Strong communication skills, with the ability to explain complex technical concepts to diverse audiences.
- Enjoy mentoring and onboarding new team members.
- Commitment to unit tests, code review, design documentation, debugging, and problem-solving.
- Deep curiosity about how systems work under the hood.
- Motivated by helping others succeed and improving system efficiency.
- Focus on improving system performance through data-informed decisions.
- Understanding of security concepts and ability to implement them to protect users and systems.
Qualifications
Bonus Points
- Passion for data visualization, graphing, and maximizing signal versus noise.
- Experience with Elasticsearch, Logstash, and Kibana.
Pay
The typical base salary range for this position is $172,500 - $260,100 annually. In select cities within the San Francisco and New York City metropolitan areas, the base salary range is $207,800 - $285,800 annually. This range represents base salary only and does not include company bonus, incentive for sales roles, equity, or benefits.
Benefits
- Time off programs
- Medical, dental, and vision coverage
- Mental health support
- Paid parental leave
- Life and disability insurance
- 401(k) retirement plan
- Employee stock purchasing program