> Markdown version of [/jobs/ext/1928722-site-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/1928722-site-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer - **Company:** Everforth Apex - **Location:** Minneapolis, MN, United States (Remote available) - **Experience:** Expert - **Salary:** $143,520.0 - $153,920.0 - **Contract:** Temporary contract - **Skills:** Java (Programming Language), .NET Framework, Artificial Intelligence, Application Performance Management, Application Release Automation, Databases, Continuous Integration, Machine Learning, Microsoft SQL Server, Oracle (Applications), Reliability Engineering, Ansible, Software Engineering, Unix Commands, Enterprise Software Applications, Large Language Models, Grafana, HybridCloud, IBM UrbanCode Deploy, Virtual Agents, Terraform, Splunk, Appdynamics, Jenkins, Artifactory - **Published:** August 5, 2026 - **Apply:** https://www.dice.com/job-detail/06bf104c-488a-457d-814c-238509439dba ## About the Role Experience: 8+ years of experience in Software Engineering, Site Reliability Engineering (SRE), Production Support, or Platform Engineering. Technical Skills: * Expertise with observability platforms (e.g., AppDynamics, Splunk, Grafana, BigPanda, Application Insights). * Experience with CI/CD tools (e.g., Jenkins, Artifactory, UDeploy, Terraform). * Automation experience using Ansible. * Experience supporting AI/ML, LLM, and Agentic AI platforms in production. * Experience with L2 level troubleshooting using Unix commands. * Familiarity with leading large-scale production support using ITIL practices. Preferred Qualifications * Experience in enterprise banking or another highly regulated industry. * Knowledge of public, hybrid, and on-prem cloud platforms. * Experience with resilient system design. * Background in supporting large-scale Java and .NET applications. * Experience with Oracle and MSSQL database support. ## Description In this contingent resource assignment, you will consult on complex initiatives with broad impact and large-scale planning for Software Engineering. This role functions as a senior Site Reliability Engineer supporting enterprise production environments, platform reliability, observability, automation, incident management, and operational excellence initiatives. The successful candidate will join a platform management organization responsible for L2/L3 production support, focusing on improving reliability, implementing automation, and maturing the team's SRE capabilities., * Provide L2/L3 support for critical enterprise applications and drive reliability improvements across production environments. * Lead incident, problem, and change management processes, including root cause analysis and corrective action planning. * Build and maintain observability solutions using platforms like Splunk, AppDynamics, Grafana, and BigPanda. * Support CI/CD pipelines and enhance release automation using tools such as Jenkins, Artifactory, UDeploy, and Terraform. * Develop operational automation and self-healing capabilities, primarily using Ansible, to reduce manual effort. * Support AI/ML, LLM, and Agentic AI-based application environments, partnering with development teams to manage reliability. * Manage operational risk and performance across on-prem, hybrid, and public cloud infrastructures. * Support large-scale Java/.NET applications and assist with troubleshooting for Oracle/MSSQL databases. ## Related Videos - [Our journey with Spring Boot in a microservice architecture](https://www.wearedevelopers.com/videos/511-our-journey-with-spring-boot-in-a-microservice-architecture) - [The Road to MLOps: How Verivox Transitioned to AWS](https://www.wearedevelopers.com/videos/1050-the-road-to-mlops-how-verivox-transitioned-to-aws) - [Dev & Test in the Cloud? Deploy your cloud environments with Ansible & Terraform](https://www.wearedevelopers.com/videos/1607-dev-test-in-the-cloud-deploy-your-cloud-environments-with-ansible-terraform) - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [Our GitOps approach for deploying an Identity Provider and an API Gateway in a SaaS company](https://www.wearedevelopers.com/videos/776-our-gitops-approach-for-deploying-an-identity-provider-and-an-api-gateway-in-a-saas-company) - [Enterprise-Cloud-Native - Fast-Paced Development & Deployment in a Highly Secure Banking Environment](https://www.wearedevelopers.com/videos/671-enterprise-cloud-native-fast-paced-development-deployment-in-a-highly-secure-banking-environment) ## Related Articles - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [Why Upskilling And Reskilling is Important For Developers](https://www.wearedevelopers.com/magazine/428-why-upskilling-and-reskilling-is-important-for-developers)