APPLICATION ADMINISTRATOR LEAD (SITE RELIABILITY ENGINEER) - JBOSS - 07212026-79334

Finance
Nashville, United States of America
yesterday

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours
Languages
English
Experience level
Senior
Compensation
$ 116K

Job location

Nashville, United States of America

Tech stack

Amazon Web Services (AWS)
Application Layers
Azure
Configuration Management
Continuous Integration
DevOps
WildFly (JBoss AS)
Performance Tuning
Reliability Engineering
Cloud Services
Ansible
Prometheus
Software Deployment
Datadog
Data Logging
Google Cloud Platform
Enterprise Software Applications
Mttr
Reliability of Systems
Infrastructure as Code (IaC)
Gitlab
Kubernetes
Puppet
Docker
Jenkins

Job description

The Application Administrator ' Lead is responsible for ensuring the reliability, availability, and performance of critical enterprise applications and infrastructure. This role supervises and leads cross-functional engineering teams, drives automation and observability initiatives, enforces operational excellence, and collaborates across IT and business units to sustain and improve service-level objectives (SLOs)., * Lead the design, automation, and operation of scalable infrastructure and application deployments.

  • Resolve complex incidents involving compute, networks, and application layers, with root cause analysis and follow-up.
  • Implement monitoring, alerting, and metrics to maintain high service availability and reduce MTTR.
  • Coordinate and automate application releases, environment migrations, and patching using CI/CD pipelines.
  • Mentor team members in engineering, DevOps practices, and tooling.
  • Enforce system reliability, security, and compliance using infrastructure-as-code and configuration management.
  • Maintain and improve runbooks, postmortems, and knowledge bases for operational continuity.
  • Collaborate with vendors and internal teams for third-party integrations, support, and lifecycle management.
  • Contribute to strategic planning with reliability-focused cost-benefit analysis and technology roadmaps.

Requirements

Education and Experience: Bachelor's degree and five years of relevant experience in system administration, infrastructure, or application support. Associate degree with equivalent experience may be substituted. Graduate coursework may replace up to two years of experience., 1. Business Insight

  1. Decision Quality
  2. Self-Development
  3. Customer Focus
  4. Instills Trust

Knowledges:

  1. Reliability Engineering & Automation
  2. Incident Response & Root Cause Analysis
  3. Performance Tuning & Scalability
  4. Infrastructure as Code (IaC)
  5. Operational Excellence

Skills:

  1. Observability (Metrics, Logging, Tracing)
  2. Communication & Cross-Team Collaboration
  3. Security & Compliance Awareness

Abilities:

  1. Perseverance
  2. Logical Thought

Tools & Equipment

  1. Observability platforms (Datadog, Prometheus)
  2. Configuration management (Ansible, Puppet)
  3. CI/CD tools (Jenkins, GitLab)
  4. Cloud services (AWS, Azure, GCP)
  5. Container orchestration (Kubernetes, Docker)

Apply for this position