Senior Reliability Engineer

Fitch Learning
London, UK
about 1 month ago
Apply on uk.indeed.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Working hours
Regular working hours
Job source

Tech stack

Agile Methodology Artificial Intelligence Amazon Elastic Compute Cloud Amazon S3 Cloud Computing Python (Programming Language) Windows PowerShell Software Deployment Software Engineering Scripting Amazon Virtual Private Cloud (VPC) Containerization
+6 more
Information Technology Performance Monitor Build Process Terraform AWS EKS Docker

Job description

Fitch Group is seeking a Reliability Engineer to join our Cloud Infrastructure and Platform Engineering team. We’re looking for someone who is curious about the evolving role of AI in infrastructure engineering, someone who actively explores how AI-assisted tooling, automation, and intelligent observability can raise the bar for reliability and developer experience. You will collaborate closely with global development and engineering teams to deliver reliable, resilient, and high-performing applications and services that drive business value.

What We Offer

  • A collaborative environment where your ideas and contributions are valued
  • Clear paths for career growth, including training and certification support
  • Access to modern tools and technologies, including emerging AI-powered engineering tool
  • Flexible work arrangements, including remote work options

What You’ll Do

  • Collaborate with globally distributed development teams to design, build, and operate reliable, resilient, and high-performing applications and services
  • Design and manage CI/CD pipelines to optimize build processes and application deployments for reliability, security, and efficiency
  • Identify, contain, and mitigate risk across all cloud environments, maintaining a robust security posture for infrastructure and applications
  • Implement proactive monitoring and observability practices to detect and prevent issues before they impact users
  • Develop and maintain automation and tooling solutions, including AI-assisted approaches to reduce toil and accelerate delivery
  • Uphold infrastructure standards and system design patterns across all cloud environments
  • Troubleshoot high-priority incidents and facilitate blameless post-mortems
  • Evaluate and adopt emerging technologies that advance our secure cloud platform strategy
  • Plan and execute disaster recovery testing to validate business continuity posture
  • Promote cost-conscious design and facilitate knowledge sharing across multidisciplinary teams
  • Participate in on-call rotation

Requirements

  • 5+ years of experience with AWS services such as VPC, EC2, ECS, EKS, ELB, S3, RDS, etc.
  • 5+ years of experience with containerization and orchestration such as Docker and Kubernetes
  • Strong knowledge with a scripting language, such as Python, Go, or PowerShell
  • Experience with modern application development practices such as CI/CD pipelines
  • Experience with Infrastructure as Code and GitOps tools such as Terraform and ArgoCD
  • Ability to clearly communicate complex technical and business concepts to team members
  • Excellent interpersonal and communication skills with the ability to effectively collaborate with multiple teams and build strong partnerships across a variety of internal and external constituencies
  • Excellent analytical and problem solving skills
  • Bachelor’s Degree or equivalent in a technology related field preferred (e.g., Computer Science, Engineering, etc.)

Nice to Have

  • Forward thinking and strategic mindset to be able to recognize and connect patterns to both enhance and simplify cloud-based development
  • Hands-on experience with GitOps for declarative, automated application delivery on AWS EKS, driving self-service deployments and reduced operational toil
  • A security first mindset along with practical experience building secure applications in the cloud
  • Practical experience building scalable cloud platforms for a large enterprise with a key drive towards developer velocity, autonomy, and automation
  • Understanding of project management tools, techniques, and methodologies such as Agile

Benefits & conditions

Actual salaries will be determined on an individualized basis and may vary based on factors including but not limited to education, training, experience, past performance, and other job-related factors. Base pay is one part of Fitch’s total compensation package, which, depending on the position, may also include commission earnings, discretionary bonuses, long-term incentives, and other benefits sponsored by Fitch.

About the company

At Fitch Group, the combined power of our global perspectives is what differentiates us. Our global network of colleagues comes together to accomplish things greater than they ever could alone.

Every team member is essential to our business and each perspective is critical to our success. We embrace a diverse culture that encourages a free exchange of ideas, guaranteeing your voice will be heard and your work will have an impact, regardless of seniority.

We are building incredible things at Fitch and we invite you to join us on our journey.

Fitch Group is a global leader in financial information services with operations in more than 30 countries. Wholly owned by the Hearst Corporation, we are comprised of three main businesses: Fitch Ratings Fitch Solutions Fitch Learning.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on uk.indeed.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

1:34 min

Essential commands for running and testing Terraform configurations

Hennie Francis · LIVE

3:49 min

Container hosting options available on Amazon Web Services

Federico Fregosi · World Congress 2022

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · World Congress 2025

1:20 min

Identifying multi-disciplinary talent for developer experience engineering roles

Hazal Mestci +1 · Coffee With Developers

2:32 min

Overview of Terraform and Terraform Cloud features

Devlin Duldulao · LIVE

Videos

See all

Related articles

See all