USNLX Ability Jobs

USNLX Ability Careers

Job Information

Cribl, Inc Senior Site Reliability Engineer (SRE) in Jackson, Mississippi

This is a Job Description for a Senior Site Reliability Engineer (SRE) in Jackson, Mississippi

Summary: Cribl Inc is seeking a Senior Site Reliability Engineer to join our mission to unlock the value of all observability data. Cribl provides users a new level of observability, intelligence, and control over their real-time data. You will join a team of technical engineers who are committed to shipping only high-quality software and enjoying all the goat gifs the internet has to offer. This role is remote and you will be part of the engineering organization where you will contribute in our efforts to envision, create, deploy, test, and ship Cribl products.

Duties & Responsibilities:

Engage with teams and improve service delivery and reliability across their entire lifecycle. Measure and monitor all production systems with an eye towards availability, latency and overall system health. Seek out the cause of errors and instability in our production cloud services and drive teams towards better operational excellence. Engage with product and platform teams to improve and evolve systems by lobbying for changes that improve reliability, resilience, and observability. Help Identify and drive down toil with creative innovation and automation. On-call responsibilities

Requirements

and Qualifications[]{#Hlk142289191}[]{#Hlk142304825}

: Extensive experience with enterprise scale continuous delivery environments. 5+ years of experience as a DevOps or SRE. Development with JavaScript/Node.js/TypeScript in a Linux/Mac environment. Experience with Configuration Management Tools like Terraform (preferred) or Puppet, Chef, Ansible. Experience with sustainable incident response in a blameless environment. Knowledge of cloud platforms (prefer Azure) and container + orchestration technologies. Experience with APM and Observability and related tools such as, New Relic, Splunk, CloudWatch, Prometheus, Grafana/Kibana, Sentry etc. Background in Linux Systems Engineering. Experience with Incident response related tools for instance, PagerDuty, Fire Hydrant, Blameless etc. Comfortable with a high level of autonomy and working with a distributed team.

Equal Opportunity/Affirmative Action Employer.

 

DirectEmployers