Site Reliability Engineer (FedRAMP / Security)
About the role
Coralogix is a modern, full-stack observability platform transforming how businesses process and understand their data. Our unique architecture powers in-stream analytics without reliance on expensive indexing or hot storage. We specialize in comprehensive monitoring of logs, metrics, trace, and security events with features such as APM, RUM, SIEM, Kubernetes monitoring, and more—all enhancing operational efficiency and reducing observability spend by up to 70%.
As a Site Reliability Engineer on our Cloud Infrastructure Team, you will focus on Enterprise FedRAMP Cloud Infrastructure, working in high-scale environments where our data pipeline processes 55TB of data daily. You will adopt cutting-edge technologies with end-to-end responsibility, build internal tools to expand platform capabilities, and collaborate with R&D to improve system stability and reliability. You will also lead the product roadmap, perform operational duties for FedRAMP cloud products (including deployments, on-call support, and incident management), and influence the direction of our product, which is designed by engineers, for engineers.
Responsibilities
- Work in high-scale environments processing large volumes of data
- Adopt cutting-edge technologies with end-to-end ownership
- Build internal tools to expand platform capabilities
- Collaborate with R&D to improve system stability and reliability
- Lead and influence the product roadmap
- Perform operational duties for FedRAMP cloud products, including deployments, on-call support, and incident management
Requirements
- At least 5 years of experience as a DevOps Engineer/SRE in production environments
- In-depth experience with Kubernetes—operating and monitoring are key
- At least 2 years of experience with FedRAMP compliance (High/Moderate levels), vulnerability management, and continuous monitoring (including scanning, patching, and reporting)—advantage
- High familiarity with monitoring tools such as Coralogix, Grafana, and Prometheus
- Experience with AWS or other cloud providers
- Experience with Infrastructure as Code (Terraform, Crossplane, etc.)
- Understanding of networking—from layers to protocols (HTTP, gRPC, SSL)
- Some software engineering experience, preferably in Golang
- Experience operating data pipelines—advantage
- Familiarity with Apache Kafka—advantage
Skills & Cultural Fit
We’re seeking candidates who are hungry, humble, and smart. Coralogix fosters a culture of innovation and continuous learning, where team members are encouraged to challenge the status quo and contribute to our shared mission. If you thrive in dynamic environments and are eager to shape the future of observability solutions, we’d love to hear from you.
Tech Stack
Kubernetes, Kops, AWS, Kafka, Prometheus, Thanos, Coralogix, Git, Argo CD, Istio, and many more.
Pay
The earnings range for this role is $170,000–$350,000, based on experience, skills, education, and work location.
Benefits
- Comprehensive healthcare, dental, and mental health benefits
- 401(k) plan with company match
- Paid sick time and paid time off