> Markdown version of [/jobs/ext/197800-site-reliability-engineer-senior-or-staff-storage-layer-services-sls](https://www.wearedevelopers.com/jobs/ext/197800-site-reliability-engineer-senior-or-staff-storage-layer-services-sls). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer (Senior or Staff), Storage Layer Services (SLS) - **Company:** MongoDB - **Location:** Findlay Township, PA, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Amazon Web Services, Microsoft Azure, Big Data, Cloud Computing, Computer Networks, Databases, Linux, Distributed Data Store, Distributed Systems, Domain Name System (DNS), Fault Tolerance, Python (Programming Language), Routing, Reliability Engineering, Software Engineering, Software Systems, TCP/IP, Transport Layer Security, Google Cloud, Multi-Cloud, Containerization, Kubernetes, Storage Technologies - **Published:** May 31, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=4a91e6c788068698 ## About the Role Do you have experience in TLS?, * Have 6+ years of experience working on software development and operating distributed systems * Proficiency in Python, Go, or a similar language * Have operated or supported stateful storage or database systems at scale, and are comfortable with durability, consistency, and recovery trade-offs. * Possess a customer-focused mindset * Value efficiency in processes and operations * Prefer automation over manual processes. We are a small team of software engineers with a strong bias towards software solutions to avoid toil * Experience using and extending containerization technologies, particularly Kubernetes, to enhance application agility, optimize resource utilization, and accelerate time-to-market * Expertise in cloud infrastructure platforms, including AWS, Google Cloud Platform (GCP), or Azure * Understanding of Linux operating system internals and networking concepts (e.g., TCP/IP, DNS, TLS, routing), * Leading major architectural shifts, such as moving from legacy storage stacks to new multi-tenant storage architectures, including planning and executing large-scale data and workload migrations with tight availability and durability requirements * Managing and scaling infrastructure across multi-cloud environments (AWS, GCP, or Azure) * Designing secure, multi-tenant runtime environments at scale ## Description * Work on our multi-tenant distributed storage systems, balancing long-term strategic infrastructure goals with immediate engineering needs * Build for reliability, making services and infrastructure available, resilient, fault-tolerant, and self-healing * Identify and configure key metrics to detect incidents and quantify service health, availability, and performance * Participate in a 24/7 on-call rotation to resolve issues involving the storage infrastructure * Become an expert in infrastructure performance, helping us optimize from the application level all the way to the kernel ## Related Videos - [An Applied Introduction to eBPF with Go](https://www.wearedevelopers.com/videos/1075-an-applied-introduction-to-ebpf-with-go) - [Reliable scalability: How Amazon.com scales on AWS](https://www.wearedevelopers.com/videos/983-reliable-scalability-how-amazon-com-scales-on-aws) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [Creating a routing app with Google Maps API from scratch](https://www.wearedevelopers.com/videos/831-creating-a-routing-app-with-google-maps-api-from-scratch) - [Turning Container security up to 11 with Capabilities](https://www.wearedevelopers.com/videos/718-turning-container-security-up-to-11-with-capabilities) - [A Technical Introduction to Bitcoin's 2nd Layer- The Lightning Network](https://www.wearedevelopers.com/videos/15-a-technical-introduction-to-bitcoin-s-2nd-layer-the-lightning-network) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [What Are The Top Skills Required For Azure Developers?](https://www.wearedevelopers.com/magazine/77-what-are-the-top-skills-required-for-azure-developers) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs)