> Markdown version of [/jobs/ext/2286319-director-of-devops-sre](https://www.wearedevelopers.com/jobs/ext/2286319-director-of-devops-sre). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Director of DevOps & SRE - **Company:** InvestorFlow - **Location:** London, UK - **Experience:** Experienced - **Salary:** £120,812.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Software System Penetration Testing, Microsoft Azure, Bash Shell, Software as a Service, Cloud Computing, Continuous Integration, Data Retention, DevOps, Disaster Recovery, Domain Name System (DNS), Github, Identity and Access Management, Python (Programming Language), Key Management, Network Security, Networking Basics, OAuth, Windows PowerShell, Redis, Prometheus, Azure DevOps Pipelines, Message Oriented Middleware, Security Assertion Markup Language (SAML), TCP/IP, Software Vulnerability Management, Azure Service Bus, Software Modules, Load Balancing, Cloud Platform System, Snowflake, Grafana, Caching, Containerization, Kubernetes, Deployment Automation, Cloudflare, Apache Kafka, Terraform, Docker, Vulnerability Analysis - **Published:** August 29, 2026 - **Apply:** https://www.adzuna.co.uk/jobs/details/5859782317 ## About the Role * 8+ years in DevOps, SRE, or infrastructure engineering, including 3+ years leading and developing engineering teams. * Deep hands-on Microsoft Azure experience at production scale: compute, data, secrets, identity services, networking and WAF, and cost/capacity management. Azure depth is essential for this role. * Strong CI/CD and release engineering background (GitHub Actions and/or Azure DevOps Pipelines) and infrastructure as code with Terraform, including module development and state management. * Production support and incident response experience for a multi-tenant SaaS platform. * Containerisation and orchestration in production (Docker, Kubernetes). * Solid grounding in identity and access management (role-based and privileged access, SSO/SAML, OAuth/Auth0) and secrets management practices. * A genuine automation-first instinct, backed by strong scripting (Python, Bash, PowerShell or similar) and a track record of replacing manual process with reliable, supportable tooling. * Networking fundamentals (TCP/IP, DNS, load balancing) and the ability to partner effectively with network and security teams. * Observability experience with Grafana, Prometheus, or equivalent, and a track record of building signal rather than noise. * Comfort operating in a security and compliance-conscious environment: financial services, private markets, or another regulated industry, including audit, penetration testing, and data retention requirements. * Excellent written and verbal communication, able to translate infrastructure risk and trade-offs for engineering peers, executives, and clients. * The judgement to know when to go deep technically and when to lead through others., Nice to have: AWS exposure alongside Azure, Snowflake or another cloud data platform, CDN and edge platforms (Cloudflare or similar), Redis, event-driven or message-oriented middleware (Kafka, Azure Service Bus), and prior experience in private equity, private credit, real assets, or capital-markets SaaS. ## Description We are looking for a Director of DevOps & SRE to lead the team responsible for the reliability, security, and scalability of the InvestorFlow platform. This is a strategic leadership role. You will set technical direction, own the DevOps/SRE roadmap, and lead a team of technical leads, senior and mid-level engineers across both disciplines, while partnering closely with Engineering and Product on technology and product roadmap planning, resourcing, and delivery sequencing. You will stay technical enough to lead architectural review, challenge design decisions, and get hands-on where the problem warrants it, but your primary leverage will come through strategy, standards, and the people you build. We take an automation-first approach to everything we do; if it is repeatable, it should be codified, and you will set that standard. You will also help shape how AI-assisted engineering and operations tooling is adopted across the organisation, an area where we are investing heavily. You Will: * Strategy and roadmap. The DevOps/SRE strategy for the InvestorFlow platform, aligned to measurable outcomes across reliability, delivery throughput, and cost. * The team. Leading, hiring, and developing technical leads, senior and mid-level engineers across DevOps and SRE, and setting priorities across infrastructure, CI/CD, security, and production support. * Cloud platform ownership. Reliability, performance, and cost-efficiency across DEV, QA, STG, and PRD -compute, data, secrets management, caching, and networking/edge security. Azure is our primary platform, with a supporting AWS and Snowflake footprint. * CI/CD and release engineering. GitHub Actions end to end, multi-region deployment automation, advanced deployment strategies (blue/green, canary, rolling) and rollback, plus security scanning, compliance checks, and vulnerability management embedded in the pipeline. * Infrastructure as code. Terraform practice and standards that standardise provisioning and reduce manual, ticket-driven infrastructure work. * Automation-first operating standard. Setting and holding the expectation that repeatable work is codified rather than performed, across provisioning, releases, access, remediation, and reporting, and driving developer enablement through self-service tooling that increases engineering capacity while maintaining controls. * Identity and access. Enterprise directory services, role-based and privileged access controls, Auth0 machine-to-machine credentials, and least-privilege access across engineering and QA. * Incident response and resilience. Escalation paths, root-cause analysis and postmortems, production readiness reviews, automated failover, and our disaster recovery and RTO/RPO commitments. You are the escalation point for high-severity incidents and major client environment changes, ensuring appropriate change management and CAB governance, with occasional off-hours support for the teams. * Observability and service levels. SLO/SLA and monitoring strategy across Grafana and our wider observability tooling, building alert coverage that is meaningful rather than noisy. * AI in engineering operations. The operational rollout of AI-assisted engineering and support tooling: automated triage, agent-based workflows, and developer-facing AI capability. * Architectural review and technical governance. Reviewing significant infrastructure and platform change with a security- and compliance-first lens, ensuring designs meet our obligations by default rather than by exception, and owning the shared-responsibility operating model between DevOps/SRE and product engineering. * Vendor and cost management. Relationships across cloud, security, CI/CD, and observability platforms, including spend and renewal negotiation. ## Related Videos - [Reducing LLM Calls with Vector Search Patterns - Raphael De Lio (Redis)](https://www.wearedevelopers.com/videos/1714-reducing-llm-calls-with-vector-search-patterns-raphael-de-lio-redis) - [Keeping applications secure by evolving OAuth 2.0 and OpenID Connect](https://www.wearedevelopers.com/videos/100152-keeping-applications-secure-by-evolving-oauth-2-0-and-openid-connect) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [The journey from developer to devops - what i've learnt along the way](https://www.wearedevelopers.com/videos/238-the-journey-from-developer-to-devops-what-i-ve-learnt-along-the-way) - [Accelerating Authentication Architecture: Taking Passwordless to the Next Level](https://www.wearedevelopers.com/videos/733-accelerating-authentication-architecture-taking-passwordless-to-the-next-level) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [What Are The Top Skills Required For Azure Developers?](https://www.wearedevelopers.com/magazine/77-what-are-the-top-skills-required-for-azure-developers) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift) - [Events like RSAC Get You CISOs. Developers Decide What Actually Gets Deployed.](https://www.wearedevelopers.com/magazine/693-events-like-rsac-get-you-cisos-developers-decide-what-actually-gets-deployed) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers)