Staff Site Reliability Engineer - OpenBlue Platform
Johnson Controls
Enhance the OpenBlue Data Platform at Johnson Controls as a Staff Site Reliability Engineer in Canada. Solve production issues while ensuring high system reliability and performance.
In this senior role, you will own escalated production challenges, applying your 7+ years of experience in site reliability and infrastructure engineering. You’ll be responsible for troubleshooting complex failures across various systems and will implement long-lasting solutions using Terraform and modern cloud technologies.
Key Responsibilities: • Serve as the senior escalation point for production incidents • Drive root cause analysis with engineering teams • Plan and execute infrastructure upgrades in Azure and AWS • Enhance observability through advanced monitoring setups • Engage directly with customers during serious incidents
Requirements: • Must reside in Canada with no sponsorship • Over 7 years in reliability engineering or similar field • Proficient with Terraform and Kubernetes in production • Familiar with Datadog and Grafana for monitoring • Willingness to perform on-call duties
Lead efforts in operational excellence and infrastructure reliability at Johnson Controls. #J-18808-Ljbffr
In this senior role, you will own escalated production challenges, applying your 7+ years of experience in site reliability and infrastructure engineering. You’ll be responsible for troubleshooting complex failures across various systems and will implement long-lasting solutions using Terraform and modern cloud technologies.
Key Responsibilities: • Serve as the senior escalation point for production incidents • Drive root cause analysis with engineering teams • Plan and execute infrastructure upgrades in Azure and AWS • Enhance observability through advanced monitoring setups • Engage directly with customers during serious incidents
Requirements: • Must reside in Canada with no sponsorship • Over 7 years in reliability engineering or similar field • Proficient with Terraform and Kubernetes in production • Familiar with Datadog and Grafana for monitoring • Willingness to perform on-call duties
Lead efforts in operational excellence and infrastructure reliability at Johnson Controls. #J-18808-Ljbffr
Reference: WJ-3875_12768058