NOC Engineer

Spectrum IT Recruitment
Birmingham, UK
5 days ago
Apply on www.reed.co.uk
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Compensation
£50,000.0
Working hours
Shift work
Job source

Tech stack

Microsoft Windows Amazon Web Services Bash Shell Cloud Computing Linux Domain Name System (DNS) Monitoring of Systems Python (Programming Language) Linux System Administration Networking Basics Reliability Engineering Cloud Services
+12 more
Prometheus TCP/IP Datadog Load Balancing Computer Network Operations Cloud Platform System Grafana Kubernetes Cloudwatch Terraform Splunk Docker

Job description

You’ll be joining an engineering-led organisation where reliability, automation and continuous improvement sit at the heart of the platform. Rather than simply responding to incidents, you’ll work to prevent them by improving systems, automating operational processes and helping shape the future of highly resilient cloud services.

If you’re passionate about building reliable cloud platforms and enjoy solving complex technical problems in large-scale production environments, we’d love to hear from you. What you’ll be doing

  • Monitoring and maintaining highly available production platforms running in AWS
  • Responding to and managing production incidents across a 24/7 service
  • Investigating complex technical issues and restoring services quickly and effectively
  • Developing automation to reduce manual operational tasks and improve platform resilience
  • Building and improving monitoring, alerting and observability across cloud environments
  • Working alongside Software, Platform, Cloud and Security Engineers to improve reliability and operational excellence
  • Contributing to post-incident reviews and driving continuous service improvements
  • Supporting containerised workloads using Kubernetes and Docker

Requirements

You’ll ideally have experience in a Production Engineering, Cloud Operations or NOC environment with exposure to:

  • Linux systems administration
  • AWS cloud infrastructure
  • Kubernetes and Docker
  • Production support and incident management
  • Python, Bash or Go scripting
  • Monitoring and observability platforms such as Grafana, Prometheus, Datadog, Splunk or CloudWatch
  • Networking fundamentals including DNS, TCP/IP and load balancing
  • A passion for automation, continuous improvement and operational excellence

Experience with Infrastructure as Code (Terraform), SRE principles (SLIs, SLOs), or regulated environments would be beneficial but isn’t essential. Why join?, * Linux

  • Windows
  • NOC
  • AWS
  • SRE
  • Site Reliability Engineer
  • Network Operations

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.reed.co.uk
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

43 sec

Software engineering journey and local Manchester roots

Jonathan Tang · Coffee With Developers

2:38 min

Establishing comprehensive monitoring and log management

Michael Eder +1 · LIVE

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · World Congress 2025

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · World Congress 2025

3:10 min

Correlating dispersed logs using structured request tracing

Michael Eder +1 · LIVE

2:34 min

Docker sandbox architecture and microVM environment integration

Manuel de la Peña Manuel de la Peña · World Congress 2026 Europe

Videos

See all

Related articles

See all