> Markdown version of [/jobs/ext/213359-site-reliability-engineer-azure](https://www.wearedevelopers.com/jobs/ext/213359-site-reliability-engineer-azure). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer, Azure - **Company:** Wellfit Technologies Inc. - **Location:** Irving, TX, United States - **Salary:** $130,000.0 - $150,000.0 - **Contract:** Permanent contract - **Skills:** .NET Framework, Application Performance Management, Application Services, Microsoft Azure, C Sharp (Programming Language), Software Debugging, Monitoring of Systems, Log Analysis, Windows PowerShell, Reliability Engineering, Prometheus, Service-Oriented Architecture, SQL Databases, Datadog, Cloud Monitoring, Grafana, Mttr, AngularJS, Information Technology, Google Cloud Functions, Dynatrace - **Published:** May 20, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=f638712070f97457 ## About the Role Do you have experience in Azure?, Do you have a Bachelor's degree?, * Bachelor's degree in Computer Science, IT, or related field. * Microsoft Azure Fundamentals (AZ-900) certification required * Proven SRE experience with a focus on monitoring, debugging, and incident response. * Extensive hands-on work with Azure App Services, Application Insights, and Azure Monitor. * Skilled with Diagnose and Troubleshoot Tools, Kudu, and PowerShell scripting. * Strong programming fundamentals with the ability to read and troubleshoot .NET/C# and Angular code. * Experience in on-call operations, incident response, and RCA writing. * Bonus: Experience with Grafana/Prometheus, DataDog/Dynatrace, Azure Front Door, CDN, Function Apps, WebJobs, Service Bus, or Event Hub. * Excellent communication, collaboration, and problem-solving skills. * Azure certifications are a strong plus. ## Description We are seeking a Site Reliability Engineer (SRE) with deep expertise in monitoring, debugging, and optimizing Azure App Services. This role is critical in ensuring our platforms remain reliable, performant, and scalable as we continue to grow. You'll combine hands-on Azure experience with code-level debugging, observability best practices, and automation to prevent issues before they occur, drive down MTTD/MTTR, and deliver an exceptional experience for patients and providers. If you thrive at the intersection of infrastructure, development, and performance, this is the role for you. What You'll Do: Monitoring & Debugging * Design, implement, and fine-tune monitoring systems for Azure-based applications. * Build custom dashboards with Azure Application Insights, Azure Monitor, and related tools. * Analyze logs, metrics, and traces to proactively troubleshoot performance and reliability issues. * Apply proficiency in C#, .NET, Angular, and SQL for code-level debugging and issue resolution. Azure App Service Expertise * Optimize application performance through a deep understanding of Azure App Service architecture. * Configure, manage, and scale App Service environments for multiple applications. Azure Tooling & Automation * Leverage Diagnose and Troubleshoot Tools, Kudu, and PowerShell scripting to resolve application and infrastructure issues. * Automate monitoring, alerting, and remediation workflows to improve reliability and reduce toil. Application Performance Monitoring * Use tools like Grafana, Prometheus, or other APM platforms to optimize system health and application performance. * Stay adaptable and quickly learn new monitoring tools and frameworks as needed. Collaboration & Communication * Partner closely with developers and operations to design effective monitoring solutions. * Document and communicate findings, solutions, and RCA reports with clarity and impact. ## Related Videos - [The OpenTelemetry mistakes I keep seeing (and how to stop making them)](https://www.wearedevelopers.com/videos/100158-the-opentelemetry-mistakes-i-keep-seeing-and-how-to-stop-making-them) - [Azure-Well Architected Framework - designing mission critical workloads in practice](https://www.wearedevelopers.com/videos/1529-azure-well-architected-framework-designing-mission-critical-workloads-in-practice) - [What Developers Get Wrong About Application Quality](https://www.wearedevelopers.com/videos/233-what-developers-get-wrong-about-application-quality) - [5 steps for running a Kubernetes environment at scale](https://www.wearedevelopers.com/videos/88-5-steps-for-running-a-kubernetes-environment-at-scale) - [All your telemetry data from any source in one place](https://www.wearedevelopers.com/videos/57-all-your-telemetry-data-from-any-source-in-one-place) - [Accelerating Authentication Architecture: Taking Passwordless to the Next Level](https://www.wearedevelopers.com/videos/733-accelerating-authentication-architecture-taking-passwordless-to-the-next-level) ## Related Articles - [What Are The Top Skills Required For Azure Developers?](https://www.wearedevelopers.com/magazine/77-what-are-the-top-skills-required-for-azure-developers) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers)