NOC Engineer
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+12 more
Job description
You’ll be joining an engineering-led organisation where reliability, automation and continuous improvement sit at the heart of the platform. Rather than simply responding to incidents, you’ll work to prevent them by improving systems, automating operational processes and helping shape the future of highly resilient cloud services.
If you’re passionate about building reliable cloud platforms and enjoy solving complex technical problems in large-scale production environments, we’d love to hear from you. What you’ll be doing
- Monitoring and maintaining highly available production platforms running in AWS
- Responding to and managing production incidents across a 24/7 service
- Investigating complex technical issues and restoring services quickly and effectively
- Developing automation to reduce manual operational tasks and improve platform resilience
- Building and improving monitoring, alerting and observability across cloud environments
- Working alongside Software, Platform, Cloud and Security Engineers to improve reliability and operational excellence
- Contributing to post-incident reviews and driving continuous service improvements
- Supporting containerised workloads using Kubernetes and Docker
Requirements
You’ll ideally have experience in a Production Engineering, Cloud Operations or NOC environment with exposure to:
- Linux systems administration
- AWS cloud infrastructure
- Kubernetes and Docker
- Production support and incident management
- Python, Bash or Go scripting
- Monitoring and observability platforms such as Grafana, Prometheus, Datadog, Splunk or CloudWatch
- Networking fundamentals including DNS, TCP/IP and load balancing
- A passion for automation, continuous improvement and operational excellence
Experience with Infrastructure as Code (Terraform), SRE principles (SLIs, SLOs), or regulated environments would be beneficial but isn’t essential. Why join?, * Linux
- Windows
- NOC
- AWS
- SRE
- Site Reliability Engineer
- Network Operations
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
Best Companies in the Netherlands: Top 25 Companies in 2023Â
Fully Remote Software Engineer Jobs
Is Software Engineering Over-Saturated?
Data Engineer Salary UK