Production Support Engineer

Sage IT Inc
Atlanta, GA, United States
1 day ago
Apply on www.dice.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
3 years minimum
Working hours
Shift work
Job source

Tech stack

Amazon Web Services Amazon Elastic Compute Cloud Amazon S3 Application Performance Management Unix Cloud Computing Databases Linux Domain Name System (DNS) Amazon DynamoDB Monitoring of Systems Tivoli Management Framework
+13 more
Lightweight Directory Access Protocols (LDAP) Simple Mail Transfer Protocols Reliability Engineering Virtualization Technology Transport Layer Security File Transfer Protocol (FTP) Load Balancing SolarWinds (Software) Route53 Cloudwatch Splunk Network Server Dynatrace

Job description

  • Manage and drive production incidents to resolution using established incident management processes.
  • Monitor and support applications deployed on AWS Cloud.
  • Troubleshoot and resolve application and infrastructure issues across AWS environments.
  • Lead technical incident triage calls and coordinate with cross-functional teams.
  • Provide timely updates on incident status, business impact, and resolution progress.
  • Analyze monitoring dashboards and identify performance trends or anomalies.
  • Perform root cause analysis (RCA) and support post-incident reviews.
  • Create and enhance operational processes, documentation, and reporting.
  • Participate in on-call rotations, including weekends and night shifts.

Requirements

  • 3+ years of experience in Production Support, Incident Management, Site Reliability Engineering (SRE), or Cloud Operations.
  • Hands-on experience with AWS services, including:
  • EC2, ELB, RDS, DynamoDB, Aurora
  • Route53, ECS, Lambda, S3
  • CloudWatch, CloudTrail, WAF, Redshift
  • Strong experience with monitoring and observability tools such as:
  • Splunk
  • Dynatrace
  • SolarWinds
  • ExtraHop
  • Catchpoint
  • MoogSoft
  • Netcool
  • Experience troubleshooting applications in AWS environments.
  • Knowledge of Linux/Unix servers, networking, DNS, LDAP, SSL, SMTP, FTP, databases, load balancers, and virtualization.
  • Ability to analyze application transactions and identify root causes across infrastructure layers.
  • Strong communication and stakeholder management skills., * AWS Solution Architect Associate Certification (or higher).
  • Experience with application performance monitoring (APM) and observability platforms.
  • Exposure to enterprise environments supporting mission-critical applications.
  • Experience preparing RCA/COE documentation and incident reporting for leadership.

Education

  • Bachelor’s Degree or equivalent experience.

Work Schedule

  • Must be willing to work in a 24x7 support environment, including on-call, weekend, and night shifts as required.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

4:23 min

Reviewing AWS infrastructure deployment configuration and planning

Devlin Duldulao · LIVE

2:38 min

Establishing comprehensive monitoring and log management

Michael Eder +1 · LIVE

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · World Congress 2025

2:03 min

Microsoft integrating native Unix coreutils into Windows environments

Chris Heilmann Chris Heilmann +2 · LIVE

3:10 min

Correlating dispersed logs using structured request tracing

Michael Eder +1 · LIVE

3:55 min

Demonstrating .NET installation on Debian and Azure Linux

Silvano Coriani Silvano Coriani · Europe 2026 Virtual

Videos

See all

Related articles

See all