Cribl - Boise, ID
posted 3 months ago
Cribl Inc is seeking a Staff Site Reliability Engineer (SRE) to join our mission to unlock the value of all observability data. As a remote-first company, we empower our employees to do their best work, wherever they are. In this role, you will be part of a team of technical engineers committed to shipping high-quality software while enjoying a collaborative and fun work environment. You will contribute to envisioning, creating, deploying, testing, and shipping Cribl products, which are designed to provide users with a new level of observability, intelligence, and control over their real-time data. As a Staff SRE, you will engage with various teams to improve service delivery and reliability across the entire lifecycle of our products. Your responsibilities will include measuring and monitoring all production systems with a focus on availability, latency, and overall system health. You will seek out the causes of errors and instability in our production cloud services and drive teams towards better operational excellence. Additionally, you will engage with product and platform teams to lobby for changes that enhance reliability, resilience, and observability. Your role will also involve identifying and driving down toil through creative innovation and automation, and you will have on-call responsibilities as part of the position. This is an exciting opportunity for those who are passionate about reliability and have strong opinions on how to improve systems. You will be involved from conception to design to development and all the way through production and beyond, providing your creative input into all things related to Cloud, Scaling, Reliability, and High Availability. If you are ready to make a real impact and be part of a team that is fundamentally changing technology, we want to hear from you!