Monitoring & Observability Engineer
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+11 more
Job description
£500 per day - 6 month contract (Initially)
About the Role
A leading global technology services organisation is seeking a skilled Monitoring & Observability Engineer to join a high-performing Professional Services team. This role focuses on designing, implementing, and managing enterprise observability solutions across complex customer environments.
You will work closely with delivery teams and customers to build scalable monitoring frameworks, improve operational visibility, and optimise application and infrastructure performance through telemetry-driven insights.
This is an excellent opportunity for an engineer with strong experience in observability platforms, cloud technologies, and automation practices who enjoys solving complex technical challenges in enterprise environments.
Key Responsibilities
- Design, deploy, configure, and maintain enterprise observability and monitoring solutions
- Implement monitoring, logging, tracing, and alerting capabilities across infrastructure and applications
- Develop and maintain observability frameworks using tools such as Dynatrace, Grafana, and Splunk
- Analyse telemetry data to identify system anomalies, performance bottlenecks, and availability risks
- Integrate observability platforms with ITSM tools and CI/CD pipelines
- Support incident response, troubleshooting, root cause analysis, and post-incident reviews
- Collaborate with internal delivery teams and customer stakeholders to deliver technical solutions
- Provide technical guidance and act as a Subject Matter Expert within monitoring and observability projects
- Identify and communicate technical risks throughout project delivery
- Contribute to continuous improvement initiatives and automation opportunities
Required Skills & Experience
- Strong hands-on experience with observability and monitoring platforms including: Dynatrace, Grafana & Splunk
- Experience collecting and analysing telemetry data including metrics, logs, traces, and events
- Strong troubleshooting and problem-solving skills within complex enterprise environments
- Experience integrating monitoring solutions with ITSM platforms such as ServiceNow
- Scripting and automation experience using Python and/or PowerShell
- Experience working with cloud platforms including Azure and AWS
- Familiarity with containerised and cloud-native environments such as Kubernetes and OpenShift
- Understanding of DevOps and CI/CD practices
- Strong communication skills with the ability to engage technical and non-technical stakeholders
- Experience within DevOps or Site Reliability Engineering (SRE) environments
- Knowledge of Infrastructure as Code tools such as Terraform
- Experience building CI/CD pipelines
- Familiarity with Docker and Kubernetes deployment practices
- Exposure to cloud-native monitoring strategies and automation frameworks
Certifications (Desirable)
- Dynatrace Associate or Professional Certification
- Splunk Core Certified Power User
- Microsoft Certified: Security Operations Analyst Associate
What We’re Looking For
- A proactive and collaborative engineer with a passion for observability and operational excellence
- Someone who can adapt quickly to new technologies and customer environments
- A methodical thinker with strong analytical and diagnostic capabilities
- A professional who values continuous learning and technical development
Apply
If you are passionate about observability engineering, cloud technologies, and enterprise monitoring solutions, we would love to hear from you.
Requirements
This is an excellent opportunity for an engineer with strong experience in observability platforms, cloud technologies, and automation practices who enjoys solving complex technical challenges in enterprise environments., * Strong hands-on experience with observability and monitoring platforms including: Dynatrace, Grafana & Splunk
- Experience collecting and analysing telemetry data including metrics, logs, traces, and events
- Strong troubleshooting and problem-solving skills within complex enterprise environments
- Experience integrating monitoring solutions with ITSM platforms such as ServiceNow
- Scripting and automation experience using Python and/or PowerShell
- Experience working with cloud platforms including Azure and AWS
- Familiarity with containerised and cloud-native environments such as Kubernetes and OpenShift
- Understanding of DevOps and CI/CD practices
- Strong communication skills with the ability to engage technical and non-technical stakeholders
- Experience within DevOps or Site Reliability Engineering (SRE) environments
- Knowledge of Infrastructure as Code tools such as Terraform
- Experience building CI/CD pipelines
- Familiarity with Docker and Kubernetes deployment practices
- Exposure to cloud-native monitoring strategies and automation frameworks, * Dynatrace Associate or Professional Certification
- Splunk Core Certified Power User
- Microsoft Certified: Security Operations Analyst Associate, * A proactive and collaborative engineer with a passion for observability and operational excellence
- Someone who can adapt quickly to new technologies and customer environments
- A methodical thinker with strong analytical and diagnostic capabilities
- A professional who values continuous learning and technical development
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
Data Engineer Salary UK
Where To Find Software Engineering Jobs
Find a Developer Job: 12 Best Job Sites For Developers
Is Software Engineering Over-Saturated?