Cribl

Staff Site Reliability Engineer

Posted Yesterday

Be an Early Applicant

Remote

Hiring Remotely in Australia

Mid level

Remote

Hiring Remotely in Australia

Mid level

As a Staff Site Reliability Engineer, you'll improve service delivery and reliability, monitor production systems, and engage with teams to enhance operational excellence. You'll require proven experience with cloud platforms, observability tools, and infrastructure management.

The summary above was generated by AI

Cribl does differently.

What does that mean? It means we are a serious company that doesn’t take itself too seriously; and we’re looking for people who love to get stuff done, and laugh a bit along the way. We’re growing rapidly - looking for collaborative, curious, and motivated team members who are passionate about putting customers first. As a remote-first company we believe in empowering our employees to do their best work, wherever they are.

As the data engine for IT and Security many of the biggest names in the most demanding industries trust Cribl to solve their most pressing data needs. Ready to do the best work of your career? Join the herd and unlock your opportunity.

Why You’ll Love This Role

Cribl Inc is seeking a Staff Site Reliability Engineer to join our mission where you will unlock the value of all observability data, as we expand our team in Australia. Cribl provides users a new level of observability, intelligence and control over their real-time data. You will join a team of technical engineers who are committed to shipping only high-quality software and enjoying all the goat gifs the internet has to offer. This role is remote, and you will be part of the engineering organization where you will contribute in our efforts to envision, create, deploy, test, and ship Cribl products.

Not often do you get to be part of something that is fundamentally changing a technology. But here at Cribl we are building the next generation of software that puts our customers in full control of their observability data. If this is something that interests you, and you want to be truly at the center of the wheel helping make this work better every day. Then this opportunity might be something you have been waiting for to be a part of making a real impact.

We are looking for Cloud Site Reliability Engineers and Developers at all levels at Cribl, who enjoy being in the thick of it. Fixing things at the operational side should always be the last resort, so our SRE engineers are involved from conception to design to development and all the way through production and beyond. You provide your creative input into all things Cloud, Scaling, Reliability, High Availability and much more.

If reliability is your passion, and you have always had strong opinions on how to make things better and have the desire to build consensus around ideas. Then let's talk!

As An Active Member Of Our Team, You Will...

Engage with teams and improve service delivery and reliability across their entire lifecycle
Measure and monitor all production systems with an eye towards availability, latency and overall system health
Seek out the cause of errors and instability in our production cloud services and drive teams towards better operational excellence
Engage with product and platform teams to improve and evolve systems by lobbying for changes that improve reliability, resilience, and observability
Help Identify and drive down toil with creative innovation and automation
This position will require stand-by, on-call, or off-hours duties

If You've Got It - We Want It

Proven experience designing, implementing, and operating observability systems for complex cloud-based platforms, with deep knowledge of best practices and a strong drive to implement them leveraging Cribl products.
Experience with Configuration Management and Infrastructure as a Code Tools like Terraform (preferred) or Ansible. Experience working with Cloud SDKs is also a plus.
Knowledge of cloud platforms (prefer AWS and Azure) and container + orchestration technologies.
Experience with APM and Observability and related tools such as, New Relic, Splunk, CloudWatch, Prometheus, Grafana/Kibana, Sentry etc.
Extensive experience with enterprise scale continuous delivery environments.
Development with JavaScript/Node.js/TypeScript in a Linux/Mac environment.
Experience with sustainable incident response in a blameless environment.
Background in Linux Systems Engineering.
Experience with Incident response related tools for instance, PagerDuty, FireHydrant, Blameless etc.
Comfortable with a high level of autonomy and working with a distributed team.
Knowledge of Cloud and application security best practices.
Strong knowledge of cloud design patterns for scale, data management, resiliency, etc.
A love for high quality and a knack for testing.
Opinions about business metrics, and SLOs.

#LI-GV1
#LI-Remote

Bring Your Whole Self
Diversity drives innovation, enables better decisions to support our customers, and inspires change for the better. We’re building a culture where differences are valued and welcomed, and we work together to bring out the best in each other. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, or any other applicable legally protected characteristics in the location in which the candidate is applying.

Interested in joining the Cribl herd? Learn more about the smartest, funniest, most passionate goats you’ll ever meet at cribl.io/about-us.

Top Skills

Ansible

AWS

Azure

Cloudwatch

Grafana

JavaScript

Kibana

Linux

New Relic

Node.js

Prometheus

Sentry

Splunk

Terraform

Typescript

Similar Jobs

CrowdStrike

Consultant

4 Hours Ago

Remote or Hybrid

Senior level

Cloud • Computer Vision • Information Technology • Sales • Security • Cybersecurity

The Principal Global Deal Review Consultant advises on Falcon Flex deals across various departments, ensuring compliance and driving revenue growth through effective deal strategies and cross-functional collaboration.

Top Skills: CpqSalesforce CRM

Boeing

Security Advisor

8 Hours Ago

Remote

Queensland, AUS

Mid level

Aerospace • Information Technology • Software • Cybersecurity • Design • Defense • Manufacturing

The Security Advisor will support Boeing Defence Australia in implementing security requirements, delivering briefings, managing administrative tasks, and supporting emergency management.

Top Skills: Asio Tech NotesDefence Security Principles Framework

Boeing

Shipping & Receiving Specialist

8 Hours Ago

Remote

Queensland, AUS

Mid level

Aerospace • Information Technology • Software • Cybersecurity • Design • Defense • Manufacturing

The Shipping & Receiving Specialist manages inventory by receipt, storage, handling, and shipping of goods, while ensuring safety and efficiency in a warehouse environment.

Top Skills: Erp SystemExcelMicrosoft OutlookMicrosoft Word

What you need to know about the Melbourne Tech Scene

Home to 650 biotech companies, 10 major research institutes and nine universities, Melbourne is among one of the top cities for biotech. In fact, some of the greatest medical advancements were conceptualized and developed here, including Symex Lab's "lab-on-a-chip" solution that monitors hormones to predict ovulation for conception, and Denteric's vaccine for periodontal gum disease. Yet, the thousands of people working in the city's healthtech sector are just getting started, to say nothing of the tech advancements across all other sectors.