Cribl - Annapolis, MD

posted 2 months ago

Full-time - Senior
Remote - Annapolis, MD

About the position

Cribl Inc is seeking a Staff Site Reliability Engineer to join our mission to unlock the value of all observability data. This role involves contributing to the engineering organization by envisioning, creating, deploying, testing, and shipping Cribl products. The position is remote-first, allowing employees to work from anywhere, and focuses on delivering high-quality software while fostering a collaborative and enjoyable work environment.

Responsibilities

  • Engage with teams to improve service delivery and reliability across their entire lifecycle.
  • Measure and monitor all production systems with an eye towards availability, latency, and overall system health.
  • Seek out the cause of errors and instability in production cloud services and drive teams towards better operational excellence.
  • Engage with product and platform teams to improve and evolve systems by lobbying for changes that enhance reliability, resilience, and observability.
  • Identify and drive down toil with creative innovation and automation.
  • Participate in on-call responsibilities.

Requirements

  • 8+ years of experience with a DevOps or SRE job title.
  • Extensive experience with enterprise scale continuous delivery environments.
  • Development experience with JavaScript/Node.js/TypeScript in a Linux/Mac environment.
  • Experience with Configuration Management Tools like Terraform (preferred), Puppet, Chef, or Ansible.
  • Knowledge of cloud platforms (preferably AWS) and container + orchestration technologies.
  • Background in Linux Systems Engineering.
  • Experience with APM and Observability tools such as New Relic, Splunk, CloudWatch, Prometheus, Grafana/Kibana, Sentry, etc.
  • Experience with incident response tools like PagerDuty, FireHydrant, Blameless, etc.
  • Comfortable with a high level of autonomy and working with a distributed team.

Nice-to-haves

  • Knowledge of Cloud and application security.
  • Strong knowledge of cloud design patterns for scale, data management, resiliency, etc.
  • A love for high quality and a knack for testing.
  • Opinions about dashboards, metrics, and SLO's.

Benefits

  • Health insurance
  • Dental insurance
  • Vision insurance
  • Short-term disability insurance
  • Life insurance
  • Paid holidays
  • Paid time off
  • Fertility treatment benefit
  • 401(k)
  • Equity
  • Eligibility for a discretionary company-wide bonus
© 2024 Teal Labs, Inc
Privacy PolicyTerms of Service