> Markdown version of [/jobs/ext/3223323-software-engineer-iii-site-reliability-engineering-ventures](https://www.wearedevelopers.com/jobs/ext/3223323-software-engineer-iii-site-reliability-engineering-ventures). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Software Engineer III, Site Reliability Engineering - Ventures - **Company:** Electronic Arts Inc. - **Location:** Guildford, UK - **Experience:** Expert - **Contract:** Temporary contract - **Skills:** Amazon Web Services, Bash Shell, Cloud Computing Security, Continuous Integration, DevOps, Github, Identity and Access Management, Python (Programming Language), Key Management, Network Configuration and Change Management, Network Segmentation, Reliability Engineering, Prometheus, Datadog, Data Logging, Grafana, Multi-Cloud, Backend, Build Management, Gitlab-ci, Cloud Optimization, Cloudwatch, Terraform, Docker, Jenkins, Golang - **Published:** September 11, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=97be2bd1faca767b ## About the Role * 5+ years of professional SRE, DevOps, platform, or infrastructure engineering experience. * Strong experience implementing observability using tools such as Datadog, Grafana, Prometheus, OpenTelemetry, or CloudWatch. * Proven experience running incident response in a production environment, including on-call and post-incident review. * Experience managing and optimising cloud cost at organisational scale, including tagging and governance practices. * Solid understanding of cloud security fundamentals, including secrets management, least-privilege access, and network segmentation. * Strong hands-on AWS experience across compute, networking, storage, IAM, and managed services. * Working knowledge of GCP, or clear evidence of operating comfortably in a multi-cloud environment. * Practical Terraform experience, sufficient to contribute confidently to an existing infrastructure as code estate. * Familiarity with CI/CD tooling such as GitHub Actions, GitLab CI, Jenkins, or similar. * Experience with containerisation and orchestration, including Docker and Kubernetes. * Strong scripting ability in Python, Bash, Go, or similar. * Experience working alongside internal platform or shared infrastructure teams within a larger organisation. * Hands-on experience using AI coding or operations agents in real work, and a clear view on where they help and where they should not be trusted. * A track record of joining an existing environment and becoming productive quickly., * Self-directed and comfortable making progress with limited ramp-up time. * Pragmatic, with good judgement about what to fix now and what to document and leave. * Genuinely enthusiastic about agentic workflows, and equally willing to say when an agent has got it wrong. * Comfortable improving systems owned by other teams without stepping on them. * Clear communicator, able to work with engineers and non-technical partners alike. * Writes things down, and leaves systems easier for others to run than they found them. * Strong problem-solving skills and attention to detail. * Commitment to inclusivity, trust, and continuous improvement. ## Description Observability and incident response (primary focus) * Design and build observability across Ventures environments, including logging, metrics, tracing, dashboards, and alerting. * Define meaningful service level indicators and alert thresholds so that teams are paged on real problems rather than noise. * Lead and improve incident response practice, including triage, escalation paths, on-call rotation, and runbooks. * Run post-incident reviews and drive the follow-up actions that come out of them. * Give product engineering teams the visibility they need to diagnose issues in their own services. Cloud cost and governance * Establish visibility of cloud spend across Ventures projects in AWS and GCP. * Identify and implement cost efficiency improvements, and set up guardrails to prevent regressions. * Define and enforce tagging, account, and resource conventions so that spend and ownership are attributable. * Report on cost trends and forecast implications for engineering leadership. Security posture * Review and improve secrets management, IAM and least-privilege access, and network configuration. * Work with EA security and platform teams to align Ventures infrastructure with internal standards. * Surface and help remediate infrastructure security risks across projects. Agentic operations * Build and refine agentic workflows that automatically inspect incidents, correlate telemetry, and propose remediations for human confirmation. * Make our systems legible to agents, including structured telemetry, machine-readable runbooks, and well-scoped tooling. * Review agent-proposed actions and remain accountable for what gets applied, treating suggestions as input rather than answers. * Track where agentic workflows perform well and where they do not, and feed that back into how they are designed. Infrastructure and delivery support * Support the backend and platform team on Terraform and infrastructure as code, contributing modules and improvements where useful. * Support and improve existing CI/CD pipelines where they create friction for product teams, without taking over ownership. * Integrate with internally provided EA infrastructure and platform services, and work with those teams to unblock delivery. * Document your work, including runbooks and operational procedures, to a standard the team can own after the contract ends., * Coverage and usefulness of monitoring and alerting across Ventures environments. * Time to detect, respond to, and resolve incidents, and the quality of post-incident follow-through. * Demonstrable reduction or better control of cloud spend, with clear attribution by project. * Measurable improvements in security posture across Ventures infrastructure. * Reduced operational friction for product engineering teams. * Growth in the coverage and accuracy of agentic workflows, and in the confidence the team has in acting on their output. * Quality of documentation and handover, measured by the team's ability to operate independently at contract end. * Effective collaboration with the backend and platform team and with EA platform and security teams. ## Related Videos - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [Go with the Flow: Stop the Leaks Before Your Memory's a Waterfall!](https://www.wearedevelopers.com/videos/100073-go-with-the-flow-stop-the-leaks-before-your-memory-s-a-waterfall) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [Shifting Stress to Progress— Understanding DevOps to do DevOps Better](https://www.wearedevelopers.com/videos/268-shifting-stress-to-progress-understanding-devops-to-do-devops-better) - [Retooling and refactoring - an investment in people.](https://www.wearedevelopers.com/videos/371-retooling-and-refactoring-an-investment-in-people) - [Designing UX for SRE Agents in High-Stakes Incidents](https://www.wearedevelopers.com/videos/100003-designing-ux-for-sre-agents-in-high-stakes-incidents) ## Related Articles - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [The 12 Best Jobs for Software Engineers](https://www.wearedevelopers.com/magazine/401-the-12-best-jobs-for-software-engineers) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Software Engineer Salary London](https://www.wearedevelopers.com/magazine/252-software-engineer-salary-london)