> Markdown version of [/jobs/ext/505979-senior-site-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/505979-senior-site-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Site Reliability Engineer - **Company:** Legora, Inc. - **Location:** New York, NY, United States - **Experience:** Expert - **Salary:** $204,000.0 - $276,000.0 - **Contract:** Permanent contract - **Skills:** Cloud Computing, Software Debugging, Fault Tolerance, Reliability Engineering, Kubernetes - **Published:** June 11, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=182a9ebc564a19b6 ## About the Role Do you have experience in Tooling?, * Significant experience operating and improving production systems, including debugging under pressure and preventing repeat incidents * Comfortable writing software and building automation to solve reliability problems * Autonomous, with pride in the quality and resilience of the systems you ship and operate * A systems thinker: failure modes, graceful degradation, and practical tradeoffs * Strong observability, incident management, and on-call experience, as well as experience with cloud infrastructure and Kubernetes ## Description As a Senior Site Reliability Engineer you'll join the founding SRE team at our new NYC engineering hub, sitting within Foundations. You'll own critical services end-to-end and partner across engineering to raise the reliability bar for the entire platform, working closely with teams in Stockholm. * This is a New York-based, 5-day in-office role. We believe building together in person drives better outcomes. What You'll Do * Build, ship, and operate foundational platform services with full ownership * Build and maintain a high-signal observability stack (metrics, logs, traces) and translate signals into action * Define and evolve SLIs/SLOs, alerting, and reliability reporting for critical systems * Improve on-call and incident response, including escalation paths, coordination, and post-incident follow-ups * Reduce toil through automation, better tooling, and improved system ergonomics * Partner with product and platform engineers to design resilient systems and improve deployment safety ## Related Videos - [Understanding Kubernetes in a visual way](https://www.wearedevelopers.com/videos/100085-understanding-kubernetes-in-a-visual-way) - [Green Cloud Computing](https://www.wearedevelopers.com/videos/592-green-cloud-computing) - [System Resilience: Surviving the Software Storm](https://www.wearedevelopers.com/videos/874-system-resilience-surviving-the-software-storm) - [Designing UX for SRE Agents in High-Stakes Incidents](https://www.wearedevelopers.com/videos/100003-designing-ux-for-sre-agents-in-high-stakes-incidents) - [Instant KAI Sandboxes with vCluster: Multi-Tenant, Multi-Scheduler GPU Sharing](https://www.wearedevelopers.com/videos/100333-instant-kai-sandboxes-with-vcluster-multi-tenant-multi-scheduler-gpu-sharing) - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [Dev Digest 131 - AI'm not sure about OSS](https://www.wearedevelopers.com/magazine/472-dev-digest-131-ai-m-not-sure-about-oss)