> Markdown version of [/jobs/ext/1408092-application-administrator-lead-site-reliability-engineer-open-shift-07212026-79339](https://www.wearedevelopers.com/jobs/ext/1408092-application-administrator-lead-site-reliability-engineer-open-shift-07212026-79339). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # APPLICATION ADMINISTRATOR LEAD (SITE RELIABILITY ENGINEER) - OPEN SHIFT - 07212026-79339 - **Company:** Finance - **Location:** Nashville, TN, United States - **Experience:** Expert - **Salary:** $89,496.0 - $116,364.0 - **Contract:** Permanent contract - **Skills:** Amazon Web Services, Application Layers, Microsoft Azure, Configuration Management, Continuous Integration, DevOps, Performance Tuning, Reliability Engineering, Cloud Services, Ansible, Prometheus, Software Deployment, Datadog, Data Logging, Google Cloud, Enterprise Software Applications, Mttr, Reliability of Systems, Infrastructure as Code (IaC), Gitlab, Kubernetes, Puppet, Docker, Jenkins - **Published:** July 23, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=33a4f5ac5352e46a ## About the Role Education and Experience: Bachelor's degree and five years of relevant experience in system administration, infrastructure, or application support. Associate degree with equivalent experience may be substituted. Graduate coursework may replace up to two years of experience., 1. Business Insight 2. Decision Quality 3. Self-Development 4. Customer Focus 5. Instills Trust Knowledges: 1. Reliability Engineering & Automation 2. Incident Response & Root Cause Analysis 3. Performance Tuning & Scalability 4. Infrastructure as Code (IaC) 5. Operational Excellence Skills: 1. Observability (Metrics, Logging, Tracing) 2. Communication & Cross-Team Collaboration 3. Security & Compliance Awareness Abilities: 1. Perseverance 2. Logical Thought Tools & Equipment 1. Observability platforms (Datadog, Prometheus) 2. Configuration management (Ansible, Puppet) 3. CI/CD tools (Jenkins, GitLab) 4. Cloud services (AWS, Azure, GCP) 5. Container orchestration (Kubernetes, Docker) ## Description The Application Administrator ' Lead is responsible for ensuring the reliability, availability, and performance of critical enterprise applications and infrastructure. This role supervises and leads cross-functional engineering teams, drives automation and observability initiatives, enforces operational excellence, and collaborates across IT and business units to sustain and improve service-level objectives (SLOs)., * Lead the design, automation, and operation of scalable infrastructure and application deployments. * Resolve complex incidents involving compute, networks, and application layers, with root cause analysis and follow-up. * Implement monitoring, alerting, and metrics to maintain high service availability and reduce MTTR. * Coordinate and automate application releases, environment migrations, and patching using CI/CD pipelines. * Mentor team members in engineering, DevOps practices, and tooling. * Enforce system reliability, security, and compliance using infrastructure-as-code and configuration management. * Maintain and improve runbooks, postmortems, and knowledge bases for operational continuity. * Collaborate with vendors and internal teams for third-party integrations, support, and lifecycle management. * Contribute to strategic planning with reliability-focused cost-benefit analysis and technology roadmaps. ## Related Videos - [What Developers Get Wrong About Application Quality](https://www.wearedevelopers.com/videos/233-what-developers-get-wrong-about-application-quality) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [Automate everything via NodeJS and Puppeteer](https://www.wearedevelopers.com/videos/322-automate-everything-via-nodejs-and-puppeteer) - [Applying Agile Principles to Incident Management ](https://www.wearedevelopers.com/videos/101-applying-agile-principles-to-incident-management) - [The Memory Leak That Ate Our Cluster: A Postmortem](https://www.wearedevelopers.com/videos/2057-the-memory-leak-that-ate-our-cluster-a-postmortem) - [Docker build without Docker](https://www.wearedevelopers.com/videos/100114-docker-build-without-docker) ## Related Articles - [What Are The Top Skills Required For Azure Developers?](https://www.wearedevelopers.com/magazine/77-what-are-the-top-skills-required-for-azure-developers) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [The Best Job Search Websites of 2025](https://www.wearedevelopers.com/magazine/368-the-best-job-search-websites-of-2025) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [DevOps Engineer Salary [2023]](https://www.wearedevelopers.com/magazine/203-devops-engineer-salary-2023)