> Markdown version of [/jobs/ext/2695481-site-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/2695481-site-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer - **Company:** Stott and May - **Location:** New York, NY, United States (Remote available) - **Salary:** $120,000.0 - $140,000.0 - **Contract:** Permanent contract - **Skills:** .NET Framework, Microsoft Windows, Artificial Intelligence, Amazon Web Services, Microsoft Azure, Bash Shell, Cloud Computing, Computer Programming, Databases, Linux, DevOps, Monitoring of Systems, Networking Basics, Octopus Deploy, Windows PowerShell, Reliability Engineering, Standard Sql, Azure Machine Learning, Software Engineering, Systems Integration, TypeScript, Datadog, Scripting, Large Language Models, Prompt Engineering, Cloudformation, Infrastructure Automation Frameworks, Performance Monitor, Teamcity, Terraform, Docker - **Published:** September 3, 2026 - **Apply:** https://find.stottandmay.com/job/site-reliability-engineer-79093/apply ## About the Role * A strong background in application support, site reliability engineering or production support. * Experience supporting production environments and resolving live incidents. * Understanding of SRE principles, including SLIs, SLOs, error budgets and blameless post-mortems. * Experience with observability and monitoring tools such as Datadog. * Experience with cloud platforms such as AWS or Azure. * Proficiency in at least one programming or scripting language, such as TypeScript, .NET, Bash or PowerShell. * Knowledge of SQL and experience working with databases. * Understanding of networking fundamentals, Linux and Windows systems. * Strong problem-solving skills and the ability to remain calm under pressure. * Excellent communication skills, with the ability to explain technical concepts to technical and non-technical stakeholders. Nice to have * Experience with infrastructure-as-code tools such as Terraform, CloudFormation or CDK. * Knowledge of DevOps deployment practices and tooling such as TeamCity, Octopus Deploy or Docker. * Experience in the financial services or fintech sector. * Exposure to AI/ML platforms, prompt engineering or integrations with LLM APIs. ## Description They are now looking for an Site Reliability Engineer to become the first technical hire in the US. This is a high-impact role for someone who enjoys solving complex production issues, improving platform reliability and working closely with engineering, infrastructure and business stakeholders. As Site Reliability Engineer, you will play a key role in ensuring the reliability, scalability and smooth operation of the company's platform and supporting infrastructure. You will contribute through incident response, proactive monitoring, automation, documentation and continuous improvement of support processes. You will also act as the on-the-ground engineering presence in the New York office, extending technical coverage into US working hours and helping bridge collaboration between US-based stakeholders and the UK engineering team. Duties and responsibilities * Investigate, own and resolve incidents across the platform and infrastructure. * Work closely with cross-functional teams to ensure a rapid and effective response to production issues. * Build and maintain runbooks, post-mortems and internal documentation to support knowledge sharing. * Improve observability through monitoring, alerting and dashboards. * Reduce operational overhead through automation and reliability-focused engineering. * Contribute to the ongoing improvement of support processes, tooling and applications. * Train, coach and support members of the UK on-call support rota. ## Related Videos - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [Docker exec without Docker](https://www.wearedevelopers.com/videos/1094-docker-exec-without-docker) - [Docker build without Docker](https://www.wearedevelopers.com/videos/100114-docker-build-without-docker) ## Related Articles - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs) - [Fullstack Developer Salary UK](https://www.wearedevelopers.com/magazine/251-fullstack-developer-salary-uk)