Sr Staff Site Reliability Engineer (SRE) - Cloud

Cribl • Canada

Company

Cribl

Location

Canada

Type

Full Time

Job Description

About Cribl

Cribl unlocks the value of observability data. Our products deliver choice and control over the rising volumes of telemetry data enabling every organization to realize the value of all their observability data without limitation. Backed by the industry’s leading venture capitalists including CRV Sequoia Capital Greylock Partners Redpoint Ventures and IVP our solutions are deployed across organizations of all sizes. Many of the biggest names in the most demanding industries trust Cribl to solve their most pressing observability needs.

At our core we foster an inclusive values-aligned culture where all belong. We believe in a remote-first operating model empowering the flexibility to do your best work wherever you are. We’re also growing rapidly welcoming collaborative curious and motivated team members who are passionate about putting customers first.

Join the herd and unlock your opportunity.

About the Team

Cribl Inc is seeking a Senior Staff Site Reliability Engineer to join our mission where you will unlock the value of all observability data as we expand our team. Cribl provides users a new level of observability intelligence and control over their real-time data. You will join a team of technical engineers who are committed to shipping only high-quality software and enjoying all the goat gifs the internet has to offer. This role is remote and you will be part of the engineering organization where you will contribute in our efforts to envision create deploy test and ship Cribl products.

Not often do you get to be part of something that is fundamentally changing a technology. But here at Cribl we are building the next generation of software that puts our customers in full control of their observability data. If this is something that interests you and you want to be truly at the center of the wheel helping make this work better every day. Then this opportunity might be something you have been waiting for to be a part of making a real impact.

We are looking for Cloud Site Reliability Engineers and Developers at all levels at Cribl who enjoy being in the thick of it. Fixing things at the operational side should always be the last resort so our SRE engineers are involved from conception to design to development and all the way through production and beyond. You provide your creative input into all things Cloud Scaling Reliability High Availability and much more.

If reliability is your passion and you have always had strong opinions on how to make things better and have the desire to build consensus around ideas. Then let's talk!

As An Active Member Of Our Team You Will...

Engage with teams and improve service delivery and reliability across their entire lifecycle
Measure and monitor all production systems with an eye towards availability latency and overall system health
Design observability systems for different types of applications using Cribl products and other OpenSource tools
Seek out the cause of errors and instability in our production cloud services and drive teams towards better operational excellence
Engage with product and platform teams to improve and evolve systems by lobbying for changes that improve reliability resilience and observability
Lead efforts enabling shift-left monitoring in the organization
Help Identify and drive down toil with creative innovation and automation
On-call responsibilities

If You Got It - We Want It

Extensive experience with enterprise-scale continuous delivery environments
Development with JavaScript/Node.js/TypeScript in a Linux/Mac environment
Experience with Configuration Management Tools like Terraform (preferred) or Puppet Chef Ansible
Knowledge of cloud platforms (prefer AWS and Azure GCP is nice to have) and container + orchestration technologies
Extensive experience designing and implementing Observability platforms based on OpenSource tools like Grafana Prometheus OpenSearch
Experience mentoring engineers and acting as Subject Matter Expert in areas of Monitoring and Observability.
Experience with native monitoring services in AWS Azure and other popular Cloud Platforms
Background in Linux Systems Engineering
Experience with Incident response tools for instance PagerDuty FireHydrant etc.
Experience with sustainable incident response in a blameless environment
Comfortable with a high level of autonomy and working with a distributed team

Preferred Qualifications

Knowledge of Cloud and application security
Strong knowledge of cloud design patterns for scale data management resiliency etc
A love for high quality and a knack for testing
Opinions about dashboards metrics and SLO’s

Bring Your Whole Self Diversity drives innovation enables better decisions to support our customers and inspires change for the better. We’re building a culture where differences are valued and welcomed. We work together to bring out the best in each other. All qualified applicants will receive consideration for employment without regard to race color religion sex sexual orientation gender identity national origin or any other applicable legally protected characteristics in the location in which the candidate is applying.

#LI-EL1

Apply Now

Date Posted

12/23/2024

Views

Back to Job Listings ❤️Add To Job List Company Info View Company Reviews

Positive

Subjectivity Score: 0.9

Similar Jobs

Intermediate Software Engineer - Athennian

Views in the last 30 days - 0

Athennian a company managing over 370000 business entities worldwide is seeking an experienced Intermediate Software Engineer The role involves design...

View Details

Staff Content Designer - Benefits & HR Apps - Gusto, Inc.

Views in the last 30 days - 0

Gusto is seeking a seasoned Content Designer to support their Benefits and HR products The role involves partnering with product teams driving custome...

View Details

Staff Software Developer - Vidyard

Views in the last 30 days - 0

Vidyard is hiring a Staff Software Developer to join their Core Team responsible for designing building and scaling the core functionality of their vi...

View Details

Clinical Data Transformation Lead - ClinChoice

Views in the last 30 days - 0

ClinChoice is seeking a Clinical Data Transformation Lead to enhance data review and cleaning processes manage data sources and ensure efficient sched...

View Details

Senior DevOps Engineer - Lemon.io

Views in the last 30 days - 0

Lemonio is a marketplace that connects Senior DevOps engineers with startups in the US and Europe They offer a monthly salary of 4k79k depending on ex...

View Details

Commercial Named Account Executive - Yubico

Views in the last 30 days - 0

Yubico founded in 2007 is a global company specializing in security keys Their mission is to make secure login easy and accessible for everyone They h...

View Details