> Markdown version of [/jobs/ext/1409452-site-reliability-engineer-sre](https://www.wearedevelopers.com/jobs/ext/1409452-site-reliability-engineer-sre). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer (SRE) - **Company:** Azure Cosmos Db - **Location:** San Francisco, CA, United States - **Experience:** Expert - **Salary:** $119,800.0 - $234,700.0 - **Contract:** Permanent contract - **Skills:** Data Analysis, Microsoft Online Services, Distributed Systems, Cloud Services, Software Engineering, Jupyter Notebook, Microsoft Power Automate, Information Technology - **Published:** July 23, 2026 - **Apply:** https://www.dice.com/job-detail/59539f38-fbad-4991-bee4-186f098eb623 ## About the Role * 6+ years technical experience in software engineering, network engineering, or systems administration + OR Bachelor's Degree in Computer Science, Information Technology, or related field AND 3+ years technical experience in software engineering, network engineering, or systems administration + OR Master's Degree in Computer Science, Information Technology, or related field AND 2+ years technical experience in software engineering, network engineering, or systems administration. Other Requirements: Ability to meet Microsoft, customer and/or government security screening requirements are required for this role. These requirements include, but are not limited to the following specialized security screenings: * + Microsoft Cloud Background Check: This position will be required to pass the Microsoft Cloud background check upon hire/transfer and every two years thereafter. Preferred/Additional Qualifications: * 4+ years of experience running large scale cloud services. * 2+ years of operational experience in improving Service Reliability, Availability and Performance. * Understanding of Observability and MELT implementation patterns for large-scale services. * Experience in Logic Apps and authoring Jupyter Notebooks. * Experience in analyzing, troubleshooting, and automating root cause analysis and mitigation of incidents impacting large-scale distributed systems. * Systematic problem-solving approach, coupled with effective communication skills and a sense of curiosity. * Ability to deal with the ambiguity associated with working in a fast-paced environment. * Influencing the product architecture and roadmap to make sure the customer-experienced supportability is always a key consideration when evolving the product. #azdat #azuredata #SRE Site Reliability Engineering IC4 - The typical base pay range for this role across the U.S. is USD $119,800 - $234,700 per year. There is a different range applicable to specific work locations, within the San Francisco Bay area and New York City metropolitan area, and the base pay range for this role in those locations is USD $160,200 - $261,000 per year. ## Description * Collaborating closely with engineering teams on building and enhancing tooling and automation solutions for faster resolution of issues impacting SLO's and averting incidents altogether when possible. * Collaborating with the customers to understand their pain points around supportability and SLO attainment and formulate strategies for addressing recurring issues in a sustainable way. * Communicate on a deeply technical level and be the single point of contact for interfacing with enterprise customers for handling service escalations and driving the issues to resolution. * Ability to design and implement any changes to service telemetry for the automation to consume if it is not already available. * Enhancing customer facing experience by proactive alerting based on utilization, trends, resource health, etc. * Analyze data and provide operational insights into customer experience to design and product teams, so that we can design features with supportability in mind. * Embody our culture and values. ## Related Videos - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [Leverage Cloud Computing Benefits with Serverless Multi-Cloud ML ](https://www.wearedevelopers.com/videos/78-leverage-cloud-computing-benefits-with-serverless-multi-cloud-ml) - [Kubernetes dev is fun, but setup and ops isn't! See a fun PaaS alternative to push any code, ipynbs or even just data!](https://www.wearedevelopers.com/videos/732-kubernetes-dev-is-fun-but-setup-and-ops-isn-t-see-a-fun-paas-alternative-to-push-any-code-ipynbs-or-even-just-data) - [Leading with Reliability: Applying SRE Principles to Build Stronger Engineering Organizations](https://www.wearedevelopers.com/videos/100185-leading-with-reliability-applying-sre-principles-to-build-stronger-engineering-organizations) - [Data Science on Software Data](https://www.wearedevelopers.com/videos/162-data-science-on-software-data) - [Introduction to Azure Machine Learning](https://www.wearedevelopers.com/videos/368-introduction-to-azure-machine-learning) ## Related Articles - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [How Much Does a Software Engineer Make? Realistic Software Engineering Salaries](https://www.wearedevelopers.com/magazine/425-how-much-does-a-software-engineer-make-realistic-software-engineering-salaries) - [Top-Paying Tech Jobs (with Salaries)](https://www.wearedevelopers.com/magazine/372-top-paying-tech-jobs-with-salaries) - [Software Engineer Salary London](https://www.wearedevelopers.com/magazine/252-software-engineer-salary-london) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [React Developer Salary [2023]](https://www.wearedevelopers.com/magazine/198-react-developer-salary-2023)