Site Reliability Engineer
Insight
Leeds, UK
28 days ago
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Apply on uk.indeed.com
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Compensation
£124,800.0 - £130,000.0
Working hours
Regular working hours
Job source
Tech stack
Amazon Web Services
Cloud Computing
Continuous Integration
Domain Name System (DNS)
Virtual Private Networks (VPN)
Python (Programming Language)
Network Troubleshooting
Network Diagnostics
Routing
Network Segmentation
Packet Analyzer
Peer-To-Peer (P2P)
+17 more
Reliability Engineering
Site Reliability Engineering Practices
Shell Script
TCP/IP
Transmission Control Protocol (TCP)
Tcpdump
Cloud-native Network Functions (CNF)
Load Balancing
Cloud Platform System
Cloud Monitoring
Istio
System Availability
Firewalls (Computer Science)
Kubernetes
Drilldown
Terraform
Dynatrace
Job description
- Troubleshoot and resolve complex GCP/AWS network incidents in production environments.
- Perform deep-dive analysis using TCP Dumps, packet captures, and network diagnostics tools.
- Diagnose and resolve issues related to TCP/IP, DNS, Routing, Firewalls, VPNs, Load Balancers, Network Overlays, Network Segregation, and Peer-to-Peer communication.
- Support and manage Kubernetes platforms (GKE/EKS).
- Build and maintain monitoring, alerting, and observability solutions.
- Perform incident response, problem management, and Root Cause Analysis (RCA).
- Automate operational tasks using Python and Shell scripting.
- Manage Infrastructure as Code (Terraform) and CI/CD pipelines.
Requirements
We are looking for a Senior Cloud Network SRE / Cloud Platform Engineer with strong expertise in GCP/AWS networking, incident management, Kubernetes, and platform reliability engineering. The ideal candidate should be capable of troubleshooting complex cloud networking issues, leading production incident resolution, and driving reliability improvements through automation and observability., * Strong hands-on experience in GCP and/or AWS Networking.
- Expertise in TCP/IP, DNS, Routing, Firewalls, VPNs, Load Balancing, and Packet Analysis.
- Experience with tcpdump, and network troubleshooting tools.
- Strong knowledge of Network Overlays, Network Segmentation, and Kubernetes Networking.
- Hands-on experience with Kubernetes (GKE/EKS).
- Experience with Dynatrace, or Cloud Monitoring tools.
- Python and Shell scripting expertise.
- Terraform and CI/CD experience.
- Strong Incident Management and RCA skills.
Good to Have
- Banking/Financial Services experience.
- Service Mesh (Istio) knowledge.
- GCP/AWS/CKA/Terraform Certifications.
- Understanding of SLI/SLO/SLA and SRE practices.
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Apply on uk.indeed.com
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
LM
Luis Minvielle
over 2 years ago
LM
Luis Minvielle
Fully Remote Software Engineer Jobs
over 2 years ago
LM
Luis Minvielle
Where To Find Software Engineering Jobs
over 2 years ago
EM
Eli McGarvie
Find a Developer Job: 12 Best Job Sites For Developers
over 3 years ago
IK
Igor Khokhriakov
How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again
23 days ago
EM
Eli McGarvie
The Best Job Search Websites of 2025
over 2 years ago