IT & Software

Staff Site Reliability Engineer - OpenBlue Platform

Johnson Controls

Richmond Hill · On · Canada

Enhance the OpenBlue Data Platform at Johnson Controls as a Staff Site Reliability Engineer in Canada. Solve production issues while ensuring high system reliability and performance.

In this senior role, you will own escalated production challenges, applying your 7+ years of experience in site reliability and infrastructure engineering. You’ll be responsible for troubleshooting complex failures across various systems and will implement long-lasting solutions using Terraform and modern cloud technologies.

Key Responsibilities: • Serve as the senior escalation point for production incidents • Drive root cause analysis with engineering teams • Plan and execute infrastructure upgrades in Azure and AWS • Enhance observability through advanced monitoring setups • Engage directly with customers during serious incidents

Requirements: • Must reside in Canada with no sponsorship • Over 7 years in reliability engineering or similar field • Proficient with Terraform and Kubernetes in production • Familiar with Datadog and Grafana for monitoring • Willingness to perform on-call duties

Lead efforts in operational excellence and infrastructure reliability at Johnson Controls. #J-18808-Ljbffr

Reference: WJ-3875_12768058

Apply now

Continue on the employer's official application - the same link they use for every candidate.

More jobs

Find more on GigBlows

This role is listed on GigBlows for discovery and search. Hiring decisions and applications are handled by the employer or their chosen application system.