Site Reliability Engineer
Key2Source INC
Charlotte, NC, United States
about 2 months ago
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Apply on dice.com
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Working hours
Regular working hours
Job source
Tech stack
Application Programming Interfaces (APIs)
Amazon Web Services
Application Integration Architecture
Microsoft Azure
Cloud Computing
Databases
Continuous Integration
DevOps
Distributed Systems
Middleware
Monitoring of Systems
Python (Programming Language)
+14 more
Windows PowerShell
Reliability Engineering
Prometheus
Scripting
Google Cloud
Enterprise Software Applications
System Availability
Grafana
Reliability of Systems
Gitlab-ci
Splunk
Appdynamics
Dynatrace
Jenkins
Job description
We are looking for an experienced Site Reliability Engineer (SRE) with strong Application Support expertise to support and enhance the reliability, stability, and performance of enterprise applications and platforms. The ideal candidate will bridge the gap between application support and reliability engineering by driving operational excellence, automation, and system resilience., * Provide L2/L3 application support for enterprise applications in production and non-production environments.
- Monitor application health, system availability, and performance using observability and monitoring tools.
- Troubleshoot and resolve application, middleware, and infrastructure-related issues within SLA timelines.
- Collaborate with Development, DevOps, Cloud, and Infrastructure teams for deployments, releases, and platform improvements.
- Perform root cause analysis (RCA) and implement permanent fixes for recurring application issues.
- Automate operational tasks, monitoring, and incident response processes.
- Support CI/CD pipelines and deployment activities across environments.
- Maintain operational documentation, runbooks, and support procedures.
- Participate in incident management, change management, and on-call support rotations.
- Continuously improve system reliability, scalability, and operational efficiency.
Requirements
- 5+ years of experience in Site Reliability Engineering and Application Support roles.
- Strong experience supporting enterprise applications in distributed environments.
- Hands-on experience with Linux/Unix systems and troubleshooting.
- Experience with monitoring and observability tools such as Splunk, Dynatrace, Grafana, AppDynamics, ELK, or Prometheus.
- Knowledge of cloud platforms such as AWS, Azure, or Google Cloud Platform.
- Experience with scripting/automation using Python, Shell, or PowerShell.
- Familiarity with CI/CD tools such as Jenkins, GitLab CI/CD, or Azure DevOps.
- Understanding of APIs, middleware, databases, and application integration troubleshooting.
- Strong analytical, problem-solving, and communication skills.
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Apply on dice.com
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
LM
Luis Minvielle
over 2 years ago
LM
Luis Minvielle
Why Upskilling And Reskilling is Important For Developers
over 2 years ago
LM
Luis Minvielle
Where To Find Software Engineering Jobs
over 2 years ago
EM
Eli McGarvie
Find a Developer Job: 12 Best Job Sites For Developers
over 3 years ago
LM
Luis Minvielle
Fully Remote Software Engineer Jobs
over 2 years ago
CH
Chris Heilmann
Dev Digest 120 - Apple and peers
about 2 years ago