NOC Engineer
Spectrum IT Recruitment
Overview
In this role you’ll help keep essential national services running by managing and improving a large-scale cloud platform. You’ll work with cross-functional teams to prevent outages, streamline operations through automation, and evolve cloud services with reliability in mind. You’ll handle 24/7 incidents, implement monitoring enhancements, and apply engineering discipline to automation and platform stability. This is a hands-on, engineering-led NOC role at a cloud-focused organization.
Pay / Benefits- bonus
- pension
- healthcare
- remote working (UK-based)
- competitive salary
- strong benefits package
- Monitor AWS production environments to ensure uptime and performance
- Own live incidents in a 24/7 rotating shift model and drive resolution
- Troubleshoot complex issues quickly to restore services
- Develop automation to reduce manual tasks and improve robustness
- Enhance monitoring, alerting, and observability tooling across the cloud estate
- Collaborate with software, platform, cloud, and security teams to raise reliability and standards
- Participate in incident retrospectives and implement improvement actions
- Maintain containerised applications using Terraform and Docker
- Background in production engineering, cloud operations, or NOC
- Linux administration experience
- Experience with AWS infrastructure
- Experience running workloads with Terraform and Docker
- Incident management and live production support
- Scripting skills in Python, Bash, or Go
- Experience with observability/monitoring tools (Grafana, Prometheus, Datadog, Splunk, CloudWatch)
- Solid understanding of DNS, TCP/IP, and load balancing
- Strong drive to automate and improve operational excellence
- Collaborative mindset
- Proactive problem-solving
- Strong ownership and accountability
- Linux administration
- AWS infrastructure
- Terraform
Reference: WJ-747_30512121