> Markdown version of [/jobs/ext/1314593-site-reliability-engineer-sre](https://www.wearedevelopers.com/jobs/ext/1314593-site-reliability-engineer-sre). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer (SRE) - **Company:** Capgemini - **Location:** Fort Mill, SC, United States - **Salary:** $76,918.0 - $120,203.0 - **Contract:** Temporary contract - **Skills:** Amazon Web Services, Application Layers, Application Performance Management, Confluence, JIRA, Microsoft Azure, Cloud Engineering, Software Documentation, DevOps, Monitoring of Systems, Octopus Deploy, Performance Tuning, Reliability Engineering, Prometheus, UML, Google Cloud, System Availability, Delivery Pipeline, Grafana, Mttr, Multi-Cloud, Infrastructure as Code (IaC), Performance Monitor, Teamcity, Terraform, Splunk, Dynatrace, Atlassian Bamboo, Jenkins - **Published:** July 17, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=1950188ca5256b87 ## About the Role We are looking for a Site Reliability Engineer with deep expertise in Dynatrace and a strong background in observability, automation, and cloud operations. This role focuses on designing and implementing highly reliable, scalable solutions while driving proactive monitoring and operational excellence., * Expert-level experience with Dynatrace, including dashboard creation, alert configuration, and integration with other observability tools. * Strong knowledge of AIOps, performance tuning, and proactive incident management. * Familiarity with hybrid/multi-cloud environments and modern DevOps practices. * Excellent problem-solving skills and ability to work in a fast-paced, collaborative environment. ## Description JIRA & Confluence aiops monitoring tools observability aws cloud Site Reliability Engineering (SRE) DevOps & CI/CD MTTR Google Cloud Platform (GCP) dynatrace, * Lead the design and implementation of full-stack observability solutions with Dynatrace as the primary platform. * Configure Dynatrace for application performance monitoring (APM), infrastructure monitoring, and intelligent alerting. * Build advanced dashboards and integrate Dynatrace with event management systems to enable proactive incident prevention and root cause analysis. * Collaborate with teams to optimize Dynatrace usage for AIOps-driven insights and automated anomaly detection. * Provide oversight for production operations to maximize reliability and automation. * Develop and evolve SRE best practices, runbooks, and tooling to ensure high availability and resilience. * Implement data-driven operational strategies to improve decision-making and reduce MTTR. * Hands-on experience with Dynatrace, Splunk, ELK, Grafana, Prometheus, and (future) ThousandEyes. * Build and manage CI/CD pipelines and Infrastructure as Code (IaC) solutions using Terraform, Jenkins, TeamCity, Octopus, Bamboo, and U-Deploy across hybrid/multi-cloud environments. * Develop and manage DevOps pipelines in AWS, Azure, and GCP using Terraform and cloud-native tooling. * Strong developer background with the ability to understand application layers and infrastructure interactions. * Define and document standard operating procedures, architecture diagrams, and system documentation using Jira, Confluence, and UML. * Identify areas for process and efficiency improvement within Platform Services Operations; recommend and implement solutions. * Drive automation initiatives across all operational processes. * Proactively monitor system capacity and health indicators; provide analytics and forecasts for scaling. ## Related Videos - [What Developers Get Wrong About Application Quality](https://www.wearedevelopers.com/videos/233-what-developers-get-wrong-about-application-quality) - [ChatGPT and Java: A Match Made in Heaven or Hell?](https://www.wearedevelopers.com/videos/536-chatgpt-and-java-a-match-made-in-heaven-or-hell) - [Improving quality with Agentic AI with Rovo Dev and Xray](https://www.wearedevelopers.com/videos/2005-improving-quality-with-agentic-ai-with-rovo-dev-and-xray) - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [Applying Agile Principles to Incident Management ](https://www.wearedevelopers.com/videos/101-applying-agile-principles-to-incident-management) - [Beyond UML: Making Sense of AI-Generated Code through Visual Architecture](https://www.wearedevelopers.com/videos/2102-beyond-uml-making-sense-of-ai-generated-code-through-visual-architecture) ## Related Articles - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [The 8 Best Code Testing Tools](https://www.wearedevelopers.com/magazine/402-the-8-best-code-testing-tools)