Sr. Network Site Reliability Engineer (SREs)

Technopride Ltd
London, UK
6 days ago
Apply on www.collegerecruiter.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
10 years minimum
Working hours
Regular working hours

Tech stack

Artificial Intelligence Amazon Web Services Microsoft Azure Border Gateway Protocol Cloud Computing Enhanced Interior Gateway Routing Protocol Virtual Private Networks (VPN) Multi-protocol Systems Python (Programming Language) Network Security Machine Learning Netconf
+20 more
Network Architecture Network Protocols Open Shortest Path First (OSPF) Overlay Transport Virtualization Performance Tuning Reliability Engineering Site Reliability Engineering Practices Ansible Zero Trust Network Access Wide Area Networks Datadog Google Cloud Grafana Software Troubleshooting Reliability of Systems Infrastructure Automation Frameworks Data Analytics Wireless Technologies Api Design Splunk

Job description

Sr. Network Site Reliability Engineer (SREs)

London, United Kingdom Posted on 09/12/2025

We provide end-to-end IT solutions and services including Applications services, Data & Analytics services, AI/ML Technologies and Professional services in the UK and EU market.

Job Description

Overview

We are seeking a highly experienced Senior Network SRE with deep expertise across multi-vendor network infrastructure, automation, and reliability engineering. The ideal candidate will possess strong technical leadership, hands-on engineering capabilities, and a passion for building resilient, scalable, and observable network environments.

Key Responsibilities

  • Design, implement, and maintain highly available network solutions across routing, switching, firewalling, and wireless technologies.
  • Apply SRE principles to improve network reliability, scalability, and performance.
  • Develop and maintain automation workflows using Ansible, Salt, and related frameworks to reduce operational toil.
  • Build and operate monitoring, alerting, and observability dashboards using tools such as Grafana and Splunk.
  • Proactively identify network bottlenecks, performance issues, and reliability risks, implementing long-term fixes rather than reactive solutions.
  • Support incident response, root cause analysis, and post-incident reviews with a focus on continuous improvement.
  • Collaborate with cross-functional engineering, security, and operations teams to ensure network solutions meet business and technical requirements.
  • Contribute to documentation, runbooks, design artifacts, and operational standards.
  • Participate in capacity planning, network modernization initiatives, and automation-first strategies.

Required Skills & Experience

  • 10+ years of hands-on experience in enterprise or service provider network engineering.
  • Expertise in multi-vendor routing, switching, firewalling, and wireless technologies.
  • Deep understanding of network protocols (BGP, OSPF, EIGRP, STP, VXLAN, VPNs, QoS, MPLS, etc.).
  • Strong experience with infrastructure automation using Ansible and Salt.
  • Proficiency with observability tooling such as Grafana, Splunk, or equivalent.
  • Solid understanding of SRE practices including SLIs, SLOs, error budgets, and proactive reliability.
  • Strong troubleshooting, analytical, and performance optimization skills.
  • Excellent communication and collaboration skills, with the ability to influence and guide technical stakeholders.

Nice to Have

  • Experience with network programmability (Python, API-driven networking, NetConf/RESTConf).
  • Exposure to cloud networking (AWS, Azure, GCP).
  • Knowledge of zero-trust, SD-WAN, and network security best practices.
  • Experience creating self-healing or fully automated network workflows.

Requirements

We are seeking a highly experienced Senior Network SRE with deep expertise across multi-vendor network infrastructure, automation, and reliability engineering. The ideal candidate will possess strong technical leadership, hands-on engineering capabilities, and a passion for building resilient, scalable, and observable network environments., * 10+ years of hands-on experience in enterprise or service provider network engineering.

  • Expertise in multi-vendor routing, switching, firewalling, and wireless technologies.
  • Deep understanding of network protocols (BGP, OSPF, EIGRP, STP, VXLAN, VPNs, QoS, MPLS, etc.).
  • Strong experience with infrastructure automation using Ansible and Salt.
  • Proficiency with observability tooling such as Grafana, Splunk, or equivalent.
  • Solid understanding of SRE practices including SLIs, SLOs, error budgets, and proactive reliability.
  • Strong troubleshooting, analytical, and performance optimization skills.
  • Excellent communication and collaboration skills, with the ability to influence and guide technical stakeholders., * Experience with network programmability (Python, API-driven networking, NetConf/RESTConf).
  • Exposure to cloud networking (AWS, Azure, GCP).
  • Knowledge of zero-trust, SD-WAN, and network security best practices.
  • Experience creating self-healing or fully automated network workflows.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.collegerecruiter.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:36 min

Visualizing memory limits and isolating suspicious endpoints

Dina Matveev Dina Matveev · Europe 2026 Virtual

10:40 min

Visualizing Prometheus open metrics using custom Grafana dashboards

Stijn Polfliet · LIVE

2:38 min

Establishing comprehensive monitoring and log management

Michael Eder +1 · LIVE

1:20 min

Identifying multi-disciplinary talent for developer experience engineering roles

Hazal Mestci +1 · Coffee With Developers

1:08 min

Analyzing error logs and root causes using artificial intelligence

Nishil Patel Nishil Patel · World Congress 2025

1:04 min

Visualizing Keycloak performance via standard Grafana troubleshooting dashboards

Alexander Schwartz Alexander Schwartz · World Congress 2025

Videos

See all

Related articles

See all