NOC Manager

Rapyd
UK
2 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Working hours
Regular working hours
Job source

Tech stack

Amazon Web Services Microsoft Azure Bash Shell Linux Python (Programming Language) Datadog Scripting Google Cloud Gitlab Kubernetes Jenkins Golang

Job description

The NOC Team Lead is responsible for ensuring the high availability, reliability, and performance of a company’s production systems, infrastructure, and applications. This hybrid role bridges traditional reactive monitoring (Network Operations Center - NOC) with proactive, automated, and software-centric engineering practices (Site Reliability Engineering - SRE)., * Operational Leadership: Directs 24/7 global NOC teams, managing incident response, service uptime, and operational excellence.

  • Proactive Monitoring & Observability: Implements monitoring, alerting, and observability tools (e.g., Prometheus, Grafana, Datadog) to track system health via golden signals (latency, traffic, errors, saturation).
  • Automation and Toil Reduction: Leads efforts to automate repetitive operational tasks and manual troubleshooting to improve system reliability and reduce human error.
  • Incident Management & Root Cause Analysis (RCA): Oversees the management of critical incidents, ensures timely communication, and performs post-mortem analysis to prevent recurrence.
  • Team Management & Development: Coaches, mentors, and develops NOC engineers and SREs, fostering a culture of high performance and continuous improvement.
  • Stakeholder Collaboration: Partners with engineering, development, and IT teams to align system performance with business goals and SLA requirements.

Key Differences in Focus

  • NOC Focus: Primarily monitoring, detecting, and responding to incidents, ensuring connectivity and managing alerts.
  • SRE Focus: Focuses on engineering improvements, reducing technical debt, automation, and system resilience.
  • The Transformation: Modern managers are transforming traditional, manual NOCs into automated, SRE-driven environments.

Requirements

  • Experience: 5+ years in managing technical teams within NOC, SRE, or infrastructure domains.
  • Technical Skills: Proficiency in cloud platforms (AWS, Azure, GCP), Kubernetes, Linux/Unix, and scripting languages (Python, Bash, Golang).
  • Tools: Experience with monitoring platforms (Datadog) and CI/CD tools (e.g., Jenkins, GitLab).
  • Soft Skills: Strong communication, leadership, and crisis management skills.

About the company

Rapyd has unified payments, payouts and fintech on one worldwide platform, and we’re assembling the world’s best team to liberate global commerce. With offices in Tel Aviv, Amsterdam, Singapore, Iceland, London, Dubai, Hong Kong, and the U.S., the opportunities at Rapyd are limitless.

We believe in straight talk, quick decisions, strong execution and elegant solutions. Rapyd is where hard work pays off and careers take off. Join us and let’s build the future of fintech together.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on rapyd.net

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:44 min

Career transition into cloud native and data management

Michael Cade · LIVE

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · WWC 2025

6:14 min

Structuring CI/CD pipelines with integrated security and quality checks

Christoph Ruggenthaler · LIVE

1:08 min

Building solutions with open source GoLang infrastructure tools

Jad Wahab · LIVE

2:39 min

Experiencing core Linux capabilities for DevOps administration

Michael Cade · LIVE

1:29 min

Summary of the Nexus RPC paradigm and use cases

Maxim Fateev Maxim Fateev · WWC 2025

Videos

See all

Related articles

See all