Monitoring & Observability Engineer

developrec
Greater London, UK
5 days ago
Apply on www.collegerecruiter.com
Prepare application

Role details

Contract type
Temporary contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours

Tech stack

Amazon Web Services Microsoft Azure Cloud Computing Cloud Engineering Cyber Security Continuous Integration DevOps Monitoring of Systems Python (Programming Language) OpenShift Windows PowerShell Reliability Engineering
+11 more
Systems Integration Data Logging Grafana Software Troubleshooting Kubernetes Infrastructure Automation Frameworks Terraform Splunk Dynatrace Docker Servicenow

Job description

£500 per day - 6 month contract (Initially)

About the Role

A leading global technology services organisation is seeking a skilled Monitoring & Observability Engineer to join a high-performing Professional Services team. This role focuses on designing, implementing, and managing enterprise observability solutions across complex customer environments.

You will work closely with delivery teams and customers to build scalable monitoring frameworks, improve operational visibility, and optimise application and infrastructure performance through telemetry-driven insights.

This is an excellent opportunity for an engineer with strong experience in observability platforms, cloud technologies, and automation practices who enjoys solving complex technical challenges in enterprise environments.

Key Responsibilities

  • Design, deploy, configure, and maintain enterprise observability and monitoring solutions
  • Implement monitoring, logging, tracing, and alerting capabilities across infrastructure and applications
  • Develop and maintain observability frameworks using tools such as Dynatrace, Grafana, and Splunk
  • Analyse telemetry data to identify system anomalies, performance bottlenecks, and availability risks
  • Integrate observability platforms with ITSM tools and CI/CD pipelines
  • Support incident response, troubleshooting, root cause analysis, and post-incident reviews
  • Collaborate with internal delivery teams and customer stakeholders to deliver technical solutions
  • Provide technical guidance and act as a Subject Matter Expert within monitoring and observability projects
  • Identify and communicate technical risks throughout project delivery
  • Contribute to continuous improvement initiatives and automation opportunities

Required Skills & Experience

  • Strong hands-on experience with observability and monitoring platforms including: Dynatrace, Grafana & Splunk
  • Experience collecting and analysing telemetry data including metrics, logs, traces, and events
  • Strong troubleshooting and problem-solving skills within complex enterprise environments
  • Experience integrating monitoring solutions with ITSM platforms such as ServiceNow
  • Scripting and automation experience using Python and/or PowerShell
  • Experience working with cloud platforms including Azure and AWS
  • Familiarity with containerised and cloud-native environments such as Kubernetes and OpenShift
  • Understanding of DevOps and CI/CD practices
  • Strong communication skills with the ability to engage technical and non-technical stakeholders
  • Experience within DevOps or Site Reliability Engineering (SRE) environments
  • Knowledge of Infrastructure as Code tools such as Terraform
  • Experience building CI/CD pipelines
  • Familiarity with Docker and Kubernetes deployment practices
  • Exposure to cloud-native monitoring strategies and automation frameworks

Certifications (Desirable)

  • Dynatrace Associate or Professional Certification
  • Splunk Core Certified Power User
  • Microsoft Certified: Security Operations Analyst Associate

What We’re Looking For

  • A proactive and collaborative engineer with a passion for observability and operational excellence
  • Someone who can adapt quickly to new technologies and customer environments
  • A methodical thinker with strong analytical and diagnostic capabilities
  • A professional who values continuous learning and technical development

Apply

If you are passionate about observability engineering, cloud technologies, and enterprise monitoring solutions, we would love to hear from you.

Requirements

This is an excellent opportunity for an engineer with strong experience in observability platforms, cloud technologies, and automation practices who enjoys solving complex technical challenges in enterprise environments., * Strong hands-on experience with observability and monitoring platforms including: Dynatrace, Grafana & Splunk

  • Experience collecting and analysing telemetry data including metrics, logs, traces, and events
  • Strong troubleshooting and problem-solving skills within complex enterprise environments
  • Experience integrating monitoring solutions with ITSM platforms such as ServiceNow
  • Scripting and automation experience using Python and/or PowerShell
  • Experience working with cloud platforms including Azure and AWS
  • Familiarity with containerised and cloud-native environments such as Kubernetes and OpenShift
  • Understanding of DevOps and CI/CD practices
  • Strong communication skills with the ability to engage technical and non-technical stakeholders
  • Experience within DevOps or Site Reliability Engineering (SRE) environments
  • Knowledge of Infrastructure as Code tools such as Terraform
  • Experience building CI/CD pipelines
  • Familiarity with Docker and Kubernetes deployment practices
  • Exposure to cloud-native monitoring strategies and automation frameworks, * Dynatrace Associate or Professional Certification
  • Splunk Core Certified Power User
  • Microsoft Certified: Security Operations Analyst Associate, * A proactive and collaborative engineer with a passion for observability and operational excellence
  • Someone who can adapt quickly to new technologies and customer environments
  • A methodical thinker with strong analytical and diagnostic capabilities
  • A professional who values continuous learning and technical development

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.collegerecruiter.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

12:33 min

Exploring advanced observability stacks and distributed infrastructure challenges

Pawel Piwosz · LIVE

2:38 min

Establishing comprehensive monitoring and log management

Michael Eder +1 · LIVE

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · World Congress 2026 Europe

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · World Congress 2025

3:10 min

Correlating dispersed logs using structured request tracing

Michael Eder +1 · LIVE

3:18 min

Scaling global network engineering through DevOps culture

Stuart Clark · LIVE

Videos

See all

Related articles

See all