> Markdown version of [/jobs/ext/3140823-azure-cloud-devops-engineer](https://www.wearedevelopers.com/jobs/ext/3140823-azure-cloud-devops-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Azure Cloud DevOps Engineer - **Company:** Tekshapers Inc - **Location:** Minneapolis, MN, United States - **Experience:** Starter - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Microsoft Azure, Bash Shell, Cloud Computing, Continuous Integration, Noise Reduction, Linux, Python (Programming Language), Pattern Recognition, Runbook, Software Vulnerability Management, Google Cloud, GitHub Copilot, Office365, Mttr, Cloudformation, Kubernetes, Terraform, Splunk, Dynatrace, Devsecops, Pagerduty, Servicenow - **Published:** September 29, 2026 - **Apply:** https://www.dice.com/job-detail/6b84fd9d-dc97-4035-94a8-627891d14f59 ## About the Role * AIOps & SRE Fundamentals: 5+ years in SRE/production operations, including SLO/SLI, error budgets, incident management, and automated remediation patterns. * Observability Toolset: 3+ years building dashboards, alerts, and troubleshooting with tools such as Dynatrace and Splunk (log/metric/trace correlation, alert tuning, noise reduction). * Cloud & Platform Engineering: 3+ years operating services on AWS/Azure/Google Cloud Platform; strong expertise in Linux, networking, containers/Kubernetes, and IaC (e.g., Terraform/CloudFormation). * Automation & Agentic Ops Mindset: Proficient in Python/Bash and CI/CD; experienced in building runbooks, self-healing workflows, and participating in on-call rotations. * AI-Enabled SRE / Intelligent Ops: 1-2+ years applying AI-assisted incident response (auto-summarization, auto-triage, pattern detection) and predictive alerting/anomaly detection in production environments. * Security / DevSecOps: Working knowledge of vulnerability management, secrets/cert governance, and secure CI/CD gates. Preferred Skills & Experience: * Release Safety & Resiliency Engineering: Experience with progressive delivery (canary/blue-green), chaos testing/DR drills, performance/capacity engineering, and well-architected reviews (e.g., Azure WARA/Azure Advisor). * ITSM & Tooling Integration: Hands-on experience integrating monitoring signals with ServiceNow and PagerDuty to build low-friction escalation workflows. General Baseline Expectation: * Enterprise AI Tool Proficiency: Demonstrate consistent use (minimum 90% weekly usage) of enterprise-approved AI tools (e.g., GitHub Copilot, Microsoft 365 Copilot) to enhance coding, documentation, and overall delivery velocity. ## Description * Own Reliability Outcomes: Define and track SLIs/SLOs, manage error budgets, and drive continuous improvement to availability, latency, and resiliency for critical services. * Operate and Mature Observability/AIOps Platform: Build and tune monitoring, dashboards, alerting, and correlation (logs/metrics/traces) using tools like Dynatrace and Splunk to reduce noise and accelerate detection/diagnosis. * Lead Incident Response & Problem Management: Run on-call/war rooms, conduct root-cause analysis, publish post-incident reviews, and ensure corrective and preventive actions are delivered. * Automate Toil & Enable Self-Healing: Create runbooks, scripts, and workflows for automated remediation, safe changes, and guardrails to improve MTTR. * Partner with Engineering & Stakeholders: Consult on architecture, release readiness, capacity planning, and operational standards; translate reliability risks into executive-ready KPI updates (MTTD/MTTR, error budget burn, recurring toil). ## Related Videos - [Our journey with Spring Boot in a microservice architecture](https://www.wearedevelopers.com/videos/511-our-journey-with-spring-boot-in-a-microservice-architecture) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [Reducing Cognitive Overload Through Platform Engineering](https://www.wearedevelopers.com/videos/679-reducing-cognitive-overload-through-platform-engineering) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [Docker exec without Docker](https://www.wearedevelopers.com/videos/1094-docker-exec-without-docker) - [Navigating the AI Wave in DevOps](https://www.wearedevelopers.com/videos/853-navigating-the-ai-wave-in-devops) ## Related Articles - [What Are The Top Skills Required For Azure Developers?](https://www.wearedevelopers.com/magazine/77-what-are-the-top-skills-required-for-azure-developers) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence](https://www.wearedevelopers.com/magazine/736-best-us-ai-conferences-for-ctos-in-2026-build-vs-buy-vendor-evaluation-and-peer-intelligence) - [What is Software Engineering in the Age of AI?](https://www.wearedevelopers.com/magazine/640-what-is-software-engineering-in-the-age-of-ai) - [7 Cloud Computing Trends Coming in 2025 for Developers](https://www.wearedevelopers.com/magazine/412-7-cloud-computing-trends-coming-in-2025-for-developers)