Cribl - Montgomery, AL
posted 3 months ago
Cribl is on a mission to unlock the value of all observability data, and we are looking for a Staff Site Reliability Engineer (SRE) to join our team. As a remote-first company, we empower our employees to do their best work from anywhere. In this role, you will be part of a collaborative and motivated team that is passionate about putting customers first. You will contribute to the development and deployment of Cribl products, ensuring high-quality software delivery while enjoying a fun and engaging work environment. As a Staff Site Reliability Engineer, you will engage with various teams to improve service delivery and reliability throughout the entire lifecycle of our systems. Your responsibilities will include measuring and monitoring production systems for availability, latency, and overall health. You will investigate the causes of errors and instability in our cloud services, driving teams towards operational excellence. Additionally, you will work closely with product and platform teams to advocate for changes that enhance reliability, resilience, and observability. This position requires a proactive approach to identifying and reducing toil through creative innovation and automation. You will also have on-call responsibilities, ensuring that our systems remain reliable and performant. If you are passionate about reliability and have strong opinions on how to improve systems, we want to hear from you!