Site Reliability Engineer with Python (IT)

Nexus
London, UK
11 days ago
Apply on www.adzuna.co.uk
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
7 years minimum
Working hours
Regular working hours

Tech stack

Java (Programming Language) JavaScript (Programming Language) Amazon Web Services Microsoft Azure Mobile Application Development Configuration Management Software Quality Databases Continuous Integration Software Debugging DevOps Github
+21 more
Issue Tracking Systems Mobile Application Software Python (Programming Language) PostgreSQL MySQL Open Source Technology Scrum Methodology Redis Reliability Engineering Software Engineering TypeScript Web Applications Circleci Git Git Flow Kubernetes Infrastructure Automation Frameworks Sumo Logic (Software) Cloudwatch Software Version Control Docker

Job description

Our Client looking to bring on a site reliability engineer to help deploy, manage, troubleshoot, and enhance our complex cloud-based set of internal tools and externally managed services for a variety of users across our wide-ranging organization.

You will have at least 7 to 10 years hands-on expertise working as a Site Reliability Engineer.

You will work closely with IT, product, and engineering to extend and maintain this set of tools and services and to help debug and resolve problems.

In addition, the ideal candidate will proactively look for system weaknesses and find ways to resolve them before they can cause production issues via monitoring and data we aggregate through various tools within our organization’s IT & DevOps toolkit.

Responsibilities

?Keep our suite of internal apps and services up and running or getting it back up and running quickly if a failure were to occur

?Be the technical point person of operational responsibility for two core platforms (one mobile and one web application) i.e. engaging as appropriate upon escalations from the IT support group whether it be problem solving, addressing production issues, enhancing features etc. - collaborating with engineers and others as needed

?Work closely with internal partners and teams as well as external vendors to ensure that we ship software that meets our code quality, security and performance requirements

?Write, update, and use our documentation, including runbooks and/or playbooks

?Help automate existing or build new internal workflows including ongoing infrastructure needs, testing, failover mitigations, and more

?Debug complex problems across our entire web and mobile application stack and advise key stakeholders on solutions, as well as implement said solutions if appropriate.

?Further our internal CI/CD processes to improve release cadence and developer experience

?Participate in the daily / weekly software development process (standups, sprint planning, retros, issue tracking, etc.)

?Actively lead any critical issue post-mortem processes, including coordination of any meetings and further steps to take

Requirements

?7+ years experience with software engineering, software development, and/or system operations

?Experience debugging complex problems and implementing timely cost-effective solutions

?Experience designing, building, and operating large-scale production systems

?Deep knowledge of Python is preferred, though other languages like Java, Go, Rust, or similar will also be heavily considered

?Experience using source control (Git, GitHub) and feature branching strategies

?Experience with a variety of open-source databases (MySQL, Postgres, Redis, etc.)

?Experience with DevOps engineering and working with container orchestration, such as with Docker or Kubernetes

?Experience with log monitoring and observability via platforms like Sumologic or Cloudwatch

?Experience automating infrastructure, testing, and deployments using tools like CircleCI Configuration management tooling and infrastructure as code knowledge is preferred but not required

?Experience working with AWS services, with knowledge of Azure / Google ecosystems helpful but not required

?Strong familiarity with general modern web and mobile application development, including hands-on experience working with JavaScript (Typescript preferred) and Python stacks

?Cross functional team collaboration experience, especially working with engineers and user experience / product designers, as well as external stakeholders

?Strong skills for weighing and managing scope, risk, quality and timelines

?Strong focus on quality, security, performance, and end user experience

This is an exciting position with an exciting organisation based in Central London and New York.

Benefits & conditions

The salary for this position will be circa £80K - £100K.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.adzuna.co.uk
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:06 min

Developer experience and project variety at scale

Alexandra Petri · World Congress 2023

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

3:55 min

Demonstrating semantic routing thresholds with the Redis vector library

2:18 min

Scaling MySQL databases for massive user growth

Johannes Nicolai Johannes Nicolai +1 · LIVE

43 sec

Software engineering journey and local Manchester roots

Jonathan Tang · Coffee With Developers

56 sec

Favorite git commands and the importance of patch commits

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

Videos

See all

Related articles

See all