Senior DevOps Engineer
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+33 more
Job description
Lumin is the technology and engineering arm of Trust Alliance Group, building the digital platforms that sit at the heart of the group’s services. From case management and dispute resolution to customer-facing digital platforms, our technology helps businesses and consumers resolve issues efficiently and build trust.We’re looking for a Senior DevOps Engineer to take ownership of our AWS estate and play a key role in shaping how we build and run technology across Lumin.This isn’t a role where you’ll simply be managing tickets or maintaining someone else’s infrastructure. You’ll have genuine ownership of our cloud environment, from Terraform and ECS to security, disaster recovery, observability, CI/CD and cost optimisation. You’ll also have the opportunity to influence how our development teams work with infrastructure, helping them become more confident and capable in owning the infrastructure behind the features they build.
What you’ll be doing:
You’ll take the lead on the infrastructure that supports three important platforms, including:
- AWS & cloud infrastructure: owning and improving our AWS estate across multiple accounts and regions.
- Infrastructure as Code: owning our Terraform estate, building reusable modules and bringing legacy infrastructure under management.
- Containers: running and improving containerised workloads on Amazon ECS, including scaling, health checks, load balancing and deployments.
- Security: owning infrastructure security across IAM, networking, secrets, encryption, credentials and vulnerability scanning.
- Disaster recovery: designing, documenting and regularly testing our multi-account, multi-region DR strategy.
- CI/CD: building and improving deployment pipelines and safe promotion through development, staging, UAT and production.
- Observability: making sure monitoring, logging and alerting tell us about problems before our users do.
- Cloud cost management: identifying and delivering opportunities to reduce AWS spend without compromising resilience or performance.
- Incident response: being a senior escalation point for infrastructure incidents and making sure root causes are properly addressed.
- Developer enablement: creating the standards, modules, tooling and review processes that allow developers to confidently own infrastructure for the features they build.
- Documentation & knowledge sharing: making sure architecture, runbooks and operational knowledge are shared across the team rather than sitting with one person.
- Continuous improvement: keeping an eye on new AWS capabilities and DevOps practices and bringing forward improvements that genuinely make us more secure, resilient, efficient or faster.
Requirements
- Significant hands-on experience running a production AWS environment at senior level.
- Strong experience with Terraform, including modules, remote state, drift, plan review and importing existing resources.
- Experience with multi-account and multi-region AWS environments.
- Experience designing, implementing and testing disaster recovery.
- Strong understanding of AWS networking, including VPCs, subnets, routing, security groups and load balancers.
- Experience running Docker/container workloads on Amazon ECS.
- Experience building and maintaining CI/CD pipelines with multiple environments and approval gates.
- Production experience with managed relational databases such as PostgreSQL, MySQL or Aurora.
- Strong understanding of infrastructure security, including IAM, least privilege, secrets and encryption.
- Experience responding to production incidents and driving root-cause analysis through to preventative action.
- Experience identifying and delivering measurable cloud cost savings.
- Experience enabling developers through infrastructure standards, code reviews, documentation or pairing.
- Experience working in an agile software development environment.
- Experience working with sensitive data and environments subject to security assurance, penetration testing or audit.
Desirable experience
It would be great if you also have experience with:
- Kubernetes / Amazon EKS.
- Amazon MQ, RabbitMQ or similar messaging technologies.
- GitHub Actions and/or CircleCI.
- Migrating source control or CI/CD platforms.
- OIDC and federated access between CI/CD and AWS.
- Serverless workloads at scale.
- Infracost, driftctl or similar cost/drift tooling.
- AI-assisted engineering with appropriate verification and guardrails.
- Java, Ruby, React, Next.js or WordPress environments.
- A second cloud platform., We’re looking for strong practical knowledge rather than someone who simply collects certifications:
- You should have hands-on experience with core AWS services including EC2, ECS, S3, RDS, Lambda, VPC, IAM, CloudWatch, Route 53, CloudFront and SSM Parameter Store / Secrets Manager.
- You’ll also need strong Terraform, Docker, AWS networking, CI/CD, scripting and observability experience. We currently use CircleCI and GitHub Actions, although equivalent experience with other tooling is relevant.
- An AWS Solutions Architect Associate certification or above, or equivalent demonstrable experience, would be expected. We support continued certification and provide protected time for professional development., * supporting production infrastructure on AWS: 3 years (required)
Benefits & conditions
- Free parking
- On-site parking
- Work from home
Application question(s):
- Have you personally owned production AWS infrastructure at senior level?
- Which Terraform activities have you personally performed in production?
- What production experience do you have with Amazon ECS?
- Have you worked with multi-account and/or multi-region AWS environments?
- Have you designed, implemented and tested production disaster recovery?
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
Fullstack Developer Salary UK
Dev Digest 121 - AI goes offline
Is Software Engineering Over-Saturated?
Fully Remote Software Engineer Jobs