> Markdown version of [/jobs/ext/2241739-lead-site-reliability-engineer-global-it](https://www.wearedevelopers.com/jobs/ext/2241739-lead-site-reliability-engineer-global-it). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Lead Site Reliability Engineer Global IT - **Company:** Fintech Farm - **Location:** Greater London, UK - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Bash Shell, Continuous Integration, DevOps, Github, Python (Programming Language), PCI Data Security Standards, Reliability Engineering, Ansible, Prometheus, Scripting, Delivery Pipeline, Grafana, Gitlab-ci, Terraform, AWS EKS, Microservices - **Published:** August 26, 2026 - **Apply:** https://www.collegerecruiter.com/job/2815093264-lead-site-reliability-engineer-global-it ## About the Role * A leader who takes ownership and inspires reliability-focused culture * Obsessed with system stability, scalability, and measurable performance * Strong communicator who can translate technical concepts into clear direction * Calm under pressure, analytical in incident response, and proactive in prevention * Passionate about mentoring engineers and driving operational excellence, * 6+ years in DevOps/SRE roles, with at least 2 years in a technical leadership position * Deep expertise inKubernetes(EKS and on-prem),Prometheus,Grafana, and alerting systems * Strong background inAWSandInfrastructure as Code(Terraform/Terragrunt) * Experience designing and maintainingCI/CD pipelines(GitLab CI/CD or GitHub Actions) * Proficiency in scripting languages (Python, Bash) and automation tooling (Ansible, Helm) * Familiar withGitOps principles(Flux, ArgoCD) * Solid understanding ofnetworking, security, and observabilitypractices * Proven ability to lead incident response and drive cross-functional reliability improvements * Exposure toDevSecOps standards, compliance, and audit processes (ISO 27001, SOC 2, PCI DSS) ## Description We are a UK fintech creating successful neobanks in emerging markets in partnerships with local traditional banks. The mission is to make banking services accessible, simple and fun to use worldwide and the goal is to launch neobanks in 50+ markets, serving 100m+ customers. Our success builds upon a best-in-class product, customer experience, emotional engagement, viral marketing and deep credit-decisioning expertise across our product suite covering credit, payments, savings and investments. One of our founders also previously co-founded a highly successful Eastern European neobank with a multi-million customer base. We launched our first market with Leobank in Azerbaijan in 2021, where we've already taken a leading market position. Our next market was Vietnam, where we launched Liobank in early 2023 and have also reached strong traction. We have several more markets on the roadmap in the next 12 months and are starting to build out teams there. Why Fintech Farm is a Great Place to Be Fintech Farm is a leading fintech with a clear mission and expansion goals. We are committed to delivering innovative banking solutions worldwide. Our Ambition We are looking to become a leading consumer digital bank brand in each market we operate, making it easy for consumers to interact with their money. You could be a part of this exciting journey. Our Culture Customers. We always go above and beyond to provide an amazing customer experience. We serve our customers the way we would want our mom to be served. And who said that banking has to be boring? We make our apps not just easy but fun to use. People. We are all business partners in our company. Each of us thinks big, acts as if we own the place and never takes 'no' for an answer. We work with strong individuals whom we empower and trust rather than micromanage. Common sense rather than formal policies prevails in all that we do. We always stay curious and open-minded. We embrace the 'we over me' culture. Your Role As a Lead SRE, you will drive the reliability, scalability, and performance of our multi-market microservices infrastructure. You'll lead a team of engineers focused on automating operations, improving observability, and ensuring zero-downtime service delivery across our cloud and on-prem environments. Your mission is to build resilient systems and empower development teams with the tools and practices needed to operate safely and efficiently at scale. What You Will Be Doing * Build and define theSRE function, establishing best practices for reliability, observability, and incident management across the platform * Manage and optimizeKubernetes clusters(AWS EKS and on-prem), ensuring scalability, cost efficiency, and resilience * Overseeobservability and alerting stack- including Prometheus, Grafana, Alertmanager, ELK, VictoriaMetrics * Implement and refine monitoring and alerting strategies, establishing actionable SLIs/SLOs and effective on-call processes * Drive improvements ininfrastructure as codeusing Terraform/Terragrunt * Collaborate closely withsoftware and DevOps teamsto ensure production readiness and reliable CI/CD delivery pipelines * Participate in and enhanceincident management processes, including post-mortems and continuous improvement initiatives * Lead efforts insecurity hardening, compliance, and cost optimizationacross environments * Contribute tostrategic planningof infrastructure roadmap and technology evolution Who You Are * A leader who takes ownership and inspires reliability-focused culture * Obsessed with system stability, scalability, and measurable performance * Strong communicator who can translate technical concepts into clear direction * Calm under pressure, analytical in incident response, and proactive in prevention * Passionate about mentoring engineers and driving operational excellence Your Experience * 6+ years in DevOps/SRE roles, with at least 2 years in a technical leadership position * Deep expertise inKubernetes(EKS and on-prem),Prometheus,Grafana, and alerting systems * Strong background inAWSandInfrastructure as Code(Terraform/Terragrunt) * Experience designing and maintainingCI/CD pipelines(GitLab CI/CD or GitHub Actions) * Proficiency in scripting languages (Python, Bash) and automation tooling (Ansible, Helm) * Familiar withGitOps principles(Flux, ArgoCD) * Solid understanding ofnetworking, security, and observabilitypractices * Proven ability to lead incident response and drive cross-functional reliability improvements * Exposure toDevSecOps standards, compliance, and audit processes (ISO 27001, SOC 2, PCI DSS) What We Are Offering * Competitive salary (negotiable based on seniority and leadership scope) * Share options * Opportunity to shape theSRE functionin a fast-scaling fintech start-up * A collaborative environment that valuesautonomy, innovation, and impact ## Related Videos - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [Dev & Test in the Cloud? Deploy your cloud environments with Ansible & Terraform](https://www.wearedevelopers.com/videos/1607-dev-test-in-the-cloud-deploy-your-cloud-environments-with-ansible-terraform) - [Innovating Developer Tools with AI: Insights from GitHub Next](https://www.wearedevelopers.com/videos/1268-innovating-developer-tools-with-ai-insights-from-github-next) - [#90DaysOfDevOps - The DevOps Learning Journey](https://www.wearedevelopers.com/videos/548-90daysofdevops-the-devops-learning-journey) - [DevOps Maturity Check – a way to balance autonomy and alignment](https://www.wearedevelopers.com/videos/58-devops-maturity-check-a-way-to-balance-autonomy-and-alignment) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Fullstack Developer Salary UK](https://www.wearedevelopers.com/magazine/251-fullstack-developer-salary-uk) - [Software Engineer Salary London](https://www.wearedevelopers.com/magazine/252-software-engineer-salary-london)