IT & Software

Cloud Reliability SRE: Incident Management & Observability

IBM

Markham · Wales · United Kingdom

IBM is seeking an expert-level Reliability Engineer to enhance system reliability within a global multi-cloud environment. This position involves analyzing failure patterns, improving tooling, and coordinating incident management practices across engineering teams. Candidates must have over 10 years of experience in SRE or incident management, strong cloud skills with AWS, GCP, or Azure, and proficiency with tools like Rootly and PagerDuty. Join IBM to shape the reliability practices that power digital transformation.
#J-18808-Ljbffr

Reference: WJ-766_18998464

Apply now

Continue on the employer's official application - the same link they use for every candidate.

More jobs

Find more on GigBlows

This role is listed on GigBlows for discovery and search. Hiring decisions and applications are handled by the employer or their chosen application system.