> Markdown version of [/jobs/ext/3290936-foundation-engineering-sre-platforms-site-reliability-engineer-associate-london](https://www.wearedevelopers.com/jobs/ext/3290936-foundation-engineering-sre-platforms-site-reliability-engineer-associate-london). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Foundation Engineering - SRE Platforms - Site Reliability Engineer - Associate - London - **Company:** Goldman Sachs, Inc. - **Location:** London, UK - **Contract:** Permanent contract - **Skills:** Java (Programming Language), Artificial Intelligence, Cloud Engineering, Computer Programming, Data Structures, Relational Databases, Software Debugging, Distributed Systems, Modular Design, Reliability Engineering, Prometheus, Software Engineering, Datadog, Data Logging, Istio, Grafana, Kubernetes, Information Technology, Microservices - **Published:** September 28, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=76f4384f7d10f6ca ## About the Role * Strong programming ability in one or more modern languages - Java or Go strongly preferred - with experience building maintainable automation beyond simple scripts. * Good understanding of networking, messaging, distributed systems, data structures, algorithms and software design fundamentals. * Hands-on experience with observability tooling, including metrics, logging, tracing and dashboarding platforms such as Prometheus, Grafana, ELK or OpenTelemetry. * Proven ability to investigate production issues, identify root causes and deliver durable engineering fixes that improve system behaviour and reduce repeat incidents. * Excellent written and verbal communication skills, with the ability to translate complex technical issues into clear updates for technical, business and senior stakeholders., * Bachelor's degree in Computer Science, Engineering or a related technical field, or equivalent practical experience. * Experience with public cloud platforms - GCP preferred - cloud-native architecture, Kubernetes, microservices or service mesh technologies. * Familiarity with relational databases and/or data-intensive platforms * Experience supporting mission-critical production systems, preferably in a financial services environment. * Experience implementing progressive delivery approaches such as canary releases, blue/green deployments, feature flags or chaos/game days. * Experience developing AI-assisted operations capabilities, such as alert enrichment, anomaly detection, triage support or runbook automation. Key SRE Competencies * Reliability engineering: SLOs, SLIs, error budgets, incident reduction and service health. * Production excellence: monitoring, alerting, observability, runbooks, handoffs and operational readiness. * Engineering mindset: automation, coding, debugging, testing, scalable design and systems thinking. * Incident leadership: calm execution, escalation discipline, clear communication and blameless learning. ## Related Videos - [Leading with Reliability: Applying SRE Principles to Build Stronger Engineering Organizations](https://www.wearedevelopers.com/videos/100185-leading-with-reliability-applying-sre-principles-to-build-stronger-engineering-organizations) - [The Memory Leak That Ate Our Cluster: A Postmortem](https://www.wearedevelopers.com/videos/2057-the-memory-leak-that-ate-our-cluster-a-postmortem) - [Rate-limiting using eBPF and Istio: How to protect your SaaS customers from themselves](https://www.wearedevelopers.com/videos/100220-rate-limiting-using-ebpf-and-istio-how-to-protect-your-saas-customers-from-themselves) - [Monitoring as Code - Managing your dashboards at scale](https://www.wearedevelopers.com/videos/753-monitoring-as-code-managing-your-dashboards-at-scale) - [Software Engineering Social Connection: Yubo’s lean approach to scaling an 80M-user infrastructure](https://www.wearedevelopers.com/videos/1583-software-engineering-social-connection-yubo-s-lean-approach-to-scaling-an-80m-user-infrastructure) - [Handling incidents collaboratively is like solving a rubix cube](https://www.wearedevelopers.com/videos/680-handling-incidents-collaboratively-is-like-solving-a-rubix-cube) ## Related Articles - [Software Engineer Salary London](https://www.wearedevelopers.com/magazine/252-software-engineer-salary-london) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Best Companies to work for in London: Top 25 Companies in 2023](https://www.wearedevelopers.com/magazine/187-best-companies-to-work-for-in-london-top-25-companies-in-2023) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs) - [How Much Does a Software Engineer Make? Realistic Software Engineering Salaries](https://www.wearedevelopers.com/magazine/425-how-much-does-a-software-engineer-make-realistic-software-engineering-salaries)