> Markdown version of [/jobs/ext/3419400-senior-site-reliability-engineer-sre](https://www.wearedevelopers.com/jobs/ext/3419400-senior-site-reliability-engineer-sre). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Site Reliability Engineer (SRE) - **Company:** Dental Intelligence - **Location:** United States (Remote available) - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Microsoft Windows, Active Directory, Amazon Web Services, Amazon Elastic Compute Cloud, Microsoft Azure, Bash Shell, Software as a Service, Cloud Computing, Cloud Engineering, Code Coverage, Code Review, Continuous Integration, DevOps, Network Topologies, Identity and Access Management, Internet Information Services (IIS), Virtual Private Networks (VPN), Python (Programming Language), Log Analysis, Windows Servers, Peering, Windows PowerShell, Systems Development Life Cycle, Role-Based Access Control, Reliability Engineering, Software Engineering, Systems Integration, Cloud Platform System, Multi-Cloud, Amazon Virtual Private Cloud (VPC), Git, Git Flow, Hashicorp, Apache Kafka, Hardware Infrastructure, Terraform - **Published:** September 10, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=30483a1036fbe4ed ## About the Role * 6+ years in SRE, DevOps, or infrastructure engineering roles, with deep, hands-on Azure experience * Strong, demonstrable Terraform expertise is required. You should be comfortable designing module structures, managing state, and using Terraform to manage complex, multi-resource environments from scratch * A genuine cloud mentality: you default to automation, reproducibility, and IaC over manual changes, and you get uncomfortable when infrastructure can't be traced back to code * A real software engineering mindset applied to infrastructure: fluency with git workflows (branching, PRs, code review), a bias toward automating anything done more than once, and discomfort with manual, undocumented changes * Experience with Windows Server administration, IIS, Active Directory/Entra ID, and Windows-based application stacks, which are critical for the on-prem component at the core of our product * Solid understanding of Azure networking (VNets, peering, private endpoints, NSGs), identity (Entra ID, RBAC, service principals and managed identities), and cost management tooling * Working knowledge of AWS core services (IAM, VPC, EC2/ECS, etc.) sufficient to support and extend an already well-architected environment * Experience integrating on-premises infrastructure with cloud environments (VPN/ExpressRoute, hybrid identity, hybrid networking) * A track record of introducing structure and standards into environments that grew organically or quickly; you enjoy imposing order, not just maintaining it * Strong scripting ability (PowerShell and/or Bash/Python) * Excellent judgment around security and access control; you think in terms of least privilege by default * Comfort operating with a high degree of autonomy and ownership, including making architectural calls and driving them through to adoption Nice to Have: * Familiarity with CI/CD tooling (Azure DevOps) for infrastructure pipelines * Prior experience leading a cloud environment from an ungoverned state to a well-architected one * Relevant certifications (Azure Solutions Architect, Azure Administrator, HashiCorp Terraform Associate) ## Description We seek an experienced Sr Site Reliability Engineer who can help our organization to mature and scale the infrastructure behind our multi-cloud SaaS platform. If the profile below sounds like you - let's talk!, We're looking for a Senior Site Reliability Engineer to help us mature and scale the infrastructure behind our multi-cloud SaaS platform. Most of our footprint runs on Microsoft Azure, built from the ground up around cloud architecture principles: autoscaling App Service and Container Apps workloads, VM Scale Sets, and Kafka-based event streaming form the backbone. We also have a smaller, well-run presence on AWS, and a Windows-based on-premises component that sits at the core of our product and integrates with our cloud environment. This role is primarily focused on Azure, where the biggest opportunity for impact lives, and you'll work across the full multi-cloud picture. We'd like to see the on-prem and cloud sides operate as a more unified, well-integrated system than they do today, and that integration work is part of what makes this role interesting. This is a high-impact role for someone who thinks in terms of systems rather than tickets, and who treats infrastructure like software. You'll have significant ownership over how our cloud infrastructure is architected, provisioned, secured, and operated going forward. If you get energized by taking a fast-growing environment and giving it real architectural rigor (consistent patterns, full infrastructure-as-code coverage, sane permission models, and cost discipline), this role was built for you. We're looking for someone who wants to build the operating model, establish the standards, and lead the transformation. You should bring genuine software engineering discipline to how that infrastructure work gets done: everything in git, everything reviewed, everything automated. No snowflakes, no manual changes made "just this once." Location: This is a fully remote role available to candidates located in U.S. states where Dental Intelligence currently has employees. What You'll do: * Own the reliability, scalability, and security posture of our Azure environment end-to-end * Lead the effort to bring our infrastructure fully under Terraform-managed IaC, replacing manual and ad-hoc provisioning with repeatable, version-controlled deployments * Treat infrastructure code like production software: everything lives in git, changes go through pull requests and peer review, and modules are tested before they ship * Define and implement a coherent Azure architecture strategy, including resource organization, naming and tagging standards, subscription and management group hierarchy, and network topology * Redesign and enforce least-privilege access and permission boundaries across Azure RBAC, Entra ID, and service principals * Identify and eliminate wasteful or redundant resource provisioning, and build cost visibility and accountability into how infrastructure is deployed * Build CI/CD pipelines for infrastructure changes so that plan/apply, validation, and policy checks are automated rather than run by hand * Build monitoring, alerting, and observability practices (Azure Monitor, Log Analytics, App Insights, or equivalent) that give the team real signal * Manage and modernize the Windows-based on-prem component that sits at the core of our product, and work to integrate it more tightly with our Azure environment * Bring the same IaC and automation discipline to bear on our AWS footprint as needed, keeping it as clean and well-run as it is today * Drive incident response, postmortems, and reliability engineering practices such as SLOs, error budgets, and capacity planning * Partner closely with engineering teams to bake reliability, security, and IaC discipline into the software delivery lifecycle * Mentor other engineers on cloud, IaC, and software engineering best practices, raising the bar across the team ## Related Videos - [How a Small Team Shrank a Microsoft Monorepo by 94%](https://www.wearedevelopers.com/videos/1236-how-a-small-team-shrank-a-microsoft-monorepo-by-94) - [Shifting Stress to Progress— Understanding DevOps to do DevOps Better](https://www.wearedevelopers.com/videos/268-shifting-stress-to-progress-understanding-devops-to-do-devops-better) - [Transforming Education: A Journey from interactive Markdown to Remote-Labs](https://www.wearedevelopers.com/videos/941-transforming-education-a-journey-from-interactive-markdown-to-remote-labs) - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [Git for Code Reviews](https://www.wearedevelopers.com/videos/429-git-for-code-reviews) - [Leading with Reliability: Applying SRE Principles to Build Stronger Engineering Organizations](https://www.wearedevelopers.com/videos/100185-leading-with-reliability-applying-sre-principles-to-build-stronger-engineering-organizations) ## Related Articles - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [What Are The Top Skills Required For Azure Developers?](https://www.wearedevelopers.com/magazine/77-what-are-the-top-skills-required-for-azure-developers) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers)