> Markdown version of [/jobs/ext/2407290-senior-platform-engineer-infrastructure](https://www.wearedevelopers.com/jobs/ext/2407290-senior-platform-engineer-infrastructure). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Platform Engineer (Infrastructure) - **Company:** Fresha - **Location:** London, UK (Remote available) - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Application Programming Interfaces (APIs), Artificial Intelligence, Amazon Web Services, Amazon Cloudfront, Amazon S3, Computing Platforms, Microsoft Azure, Bash Shell, Computer Programming, Linux, Domain Name System (DNS), Github, Monitoring of Systems, Virtual Private Networks (VPN), Python (Programming Language), PostgreSQL, Networking Basics, Operational Data Store, Redis, Ruby, Next.js, Message Oriented Middleware, TCP/IP, TypeScript, Datadog, Load Balancing, Large Language Models, Grafana, Firewalls (Computer Science), Amazon Relational Database Service, Sentry, Apache Kafka, Graphql, Front End Software Development, Terraform, AWS EKS, Docker, Elixir, Pagerduty, Jenkins, Golang - **Published:** August 27, 2026 - **Apply:** https://uk.indeed.com/viewjob?jk=7c7f17bfda881931 ## About the Role * Have 5+ years of experience building, operating, and troubleshooting production-grade systems, with a strong focus on reliability, observability, and scalability * Bring hands-on experience with AI/ML or LLM-based tools in a platform or operational context (e.g. automation, observability, developer experience, or incident response) * Demonstrate strong programming and automation skills in Python, Go, or Bash, and use Infrastructure as Code (Terraform) to build repeatable, low-risk systems * Have experience with AWS (EKS, RDS, S3, CloudFront), or equivalent platforms on GCP or Azure * Are comfortable working close to the system with solid Linux and networking fundamentals (TCP/IP, DNS, firewalls, load balancing, VPNs) * Have practical experience designing or improving observability (metrics, logs, traces) using tools such as Datadog, Grafana, ELK, Sentry, and OpsGenie * Take ownership of complex technical problems, form clear opinions backed by data, and drive solutions through implementation and collaboration * Work effectively across teams, communicate technical concepts clearly, and apply systems thinking to make platforms easier to use, safer to operate, and simpler to scale * Have experience with Claude code ## Description * Docker * AWS EKS * Kafka / AutoMQ for asynchronous messaging * Elixir & Ruby for core services * gRPC for inter-service communication * GraphQL for API ingress * Next.js / TypeScript for frontend * PostgreSQL (RDS) for persistent storage * GitHub Actions (primary CI), with Jenkins & Argo Workflows for legacy pipelines * Redis for key/value storage * Terraform for infrastructure as code * Python and GoLang for our Platform tools * Datadog, Sentry for observability and incident response What you will be doing: Platform Architecture & Scale * Define and evolve infrastructure architecture to support multi-region deployments * Design systems for resilience, scalability, and operational simplicity * Explore and apply AI-assisted approaches to reduce operational toil, improve reliability, and support platform decision-making where it creates clear value Observability & Operations * Extend monitoring and observability capabilities across services * Make operational data easy to access, understand, and act on * Build tooling for safe deployments, fast rollbacks, and reduced operational toil using software methodologies and approaches * Experiment with AI-assisted detection, triage, or automation to improve signal quality and reduce manual effort Ownership & Leadership * Scope, lead, and deliver platform initiatives autonomously * Drive cross-team projects with clear outcomes and accountability Enablement & Knowledge Sharing * Raise the bar through documentation, runbooks, and internal knowledge sharing * Help teams learn, align, and operate more effectively together ## Related Videos - [Coffee with Developers: David Heinemeier Hansson](https://www.wearedevelopers.com/videos/875-coffee-with-developers-david-heinemeier-hansson) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [Reducing LLM Calls with Vector Search Patterns - Raphael De Lio (Redis)](https://www.wearedevelopers.com/videos/1714-reducing-llm-calls-with-vector-search-patterns-raphael-de-lio-redis) - [Platform Engineering vs. DevOps Why not both?](https://www.wearedevelopers.com/videos/885-platform-engineering-vs-devops-why-not-both) - [Coroutine explained yet again 60 years later](https://www.wearedevelopers.com/videos/690-coroutine-explained-yet-again-60-years-later) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) ## Related Articles - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Why Upskilling And Reskilling is Important For Developers](https://www.wearedevelopers.com/magazine/428-why-upskilling-and-reskilling-is-important-for-developers) - [Where to Find Entry-Level Software Engineering Jobs](https://www.wearedevelopers.com/magazine/397-where-to-find-entry-level-software-engineering-jobs) - [7 Most Popular Web Developer Jobs in Europe](https://www.wearedevelopers.com/magazine/163-7-most-popular-web-developer-jobs-in-europe)