> Markdown version of [/jobs/ext/1415221-site-reliability-engineer-london](https://www.wearedevelopers.com/jobs/ext/1415221-site-reliability-engineer-london). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer- London - **Company:** FDM Group - **Location:** London, UK - **Contract:** Temporary contract - **Skills:** Amazon Web Services, Data Analysis, Microsoft Azure, Backup Devices, Bash Shell, DevOps, Elasticsearch, Monitoring of Systems, Python (Programming Language), Performance Tuning, Windows PowerShell, Reliability Engineering, Data Logging, Scripting, Grafana, Database Optimization, Containerization, Kubernetes, Docker - **Published:** July 24, 2026 - **Apply:** https://uk.indeed.com/viewjob?jk=eb45a7c1d184d4a1 ## About the Role * Extensive hands-on experience deploying and managing OpenSearch in production environments. * Deep understanding of OpenSearch architecture, cluster design, indexing strategies, shard management, and performance tuning. * Experience implementing log aggregation, search, analytics, and observability use cases using OpenSearch. * Knowledge of OpenSearch security, access controls, backups, upgrades, and operational best practices. Monitoring & Observability * Strong experience with Grafana, including dashboard development, alerting, and data source integration. * Experience with enterprise monitoring platforms, specifically Geneos. * Understanding of modern observability principles, including metrics, logs, traces, alerting, and service health monitoring. Scripting & Automation * Strong scripting skills in one or more of: + Python + Shell/Bash + PowerShell * Experience automating operational tasks and monitoring workflows. SRE / Platform Engineering * Proven experience in an SRE, Platform Engineering, DevOps, or Infrastructure Engineering role. * Strong troubleshooting and problem-solving capabilities. * Experience supporting highly available and business-critical systems. * Understanding of incident management, resilience engineering, and operational excellence practices. Desirable Skills * Experience with cloud platforms (Azure, AWS, or GCP). * Knowledge of containerisation technologies (Docker, Kubernetes). * Experience with CI/CD pipelines and Infrastructure as Code. * Experience working within financial services or regulated environments. * Familiarity with Elasticsearch ecosystems and migration strategies to OpenSearch., The ideal candidate will be a hands-on engineer who combines deep technical expertise with a pragmatic operational mindset. They will be comfortable working independently, driving observability improvements, and collaborating across engineering teams to deliver reliable and scalable monitoring solutions. Key attributes: * Strong ownership mentality. * Excellent analytical and troubleshooting skills. * Effective stakeholder communication. * Ability to operate in fast-paced production environments. * Focus on reliability, automation, and continuous improvements ## Description FDM is a global business and technology consultancy seeking a Site Reliability Engineer to work for our client within the Finance sector. This is initially a 6 month contract with very good prospects to extend and will be a hybrid role that will be based in London. Our client is seeking an experienced Site Reliability Engineer (SRE) with a strong focus on Observability and Monitoring Platforms. The successful candidate will play a key role in enhancing the organisation's monitoring, alerting, and operational visibility capabilities across critical engineering systems. This role requires hands-on expertise in the deployment, administration, and optimisation of OpenSearch, alongside experience with Grafana, Geneos, and automation/scripting technologies. Particular emphasis will be placed on the candidate's ability to design, deploy, and support enterprise-grade OpenSearch environments. Responsibilities: * Lead the design, deployment, configuration, and ongoing management of OpenSearch clusters and associated observability tooling. * Develop and maintain scalable monitoring, logging, and alerting solutions for business-critical applications and infrastructure. * Build and enhance observability dashboards using Grafana. * Support and optimise existing Geneos monitoring implementations. * Create and maintain automation scripts to streamline operational processes and improve reliability. * Collaborate with engineering, infrastructure, and support teams to improve system resilience and operational performance. * Define and implement SRE best practices, including monitoring standards, alert management, incident response, and operational readiness. * Perform troubleshooting and root cause analysis of platform and application issues. * Support capacity planning, performance tuning, and platform optimisation initiatives. * Contribute to documentation, operational procedures, and knowledge sharing within the engineering team. ## Related Videos - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [5 steps for running a Kubernetes environment at scale](https://www.wearedevelopers.com/videos/88-5-steps-for-running-a-kubernetes-environment-at-scale) - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [#90DaysOfDevOps - The DevOps Learning Journey](https://www.wearedevelopers.com/videos/548-90daysofdevops-the-devops-learning-journey) - [All your telemetry data from any source in one place](https://www.wearedevelopers.com/videos/57-all-your-telemetry-data-from-any-source-in-one-place) ## Related Articles - [Best Companies to work for in London: Top 25 Companies in 2023](https://www.wearedevelopers.com/magazine/187-best-companies-to-work-for-in-london-top-25-companies-in-2023) - [Software Engineer Salary London](https://www.wearedevelopers.com/magazine/252-software-engineer-salary-london) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk)