Site Reliability Engineer 2

Oracle
Nashville, TN, United States
about 1 month ago
Apply on eeho.fa.us2.oraclecloud.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
3 years minimum
Working hours
Regular working hours

Tech stack

Cerner Microsoft Windows Artificial Intelligence Applications Architecture Bash Shell Cloud Computing Configuration Management Cyber Security Computer Networks System Configuration Linux DevOps
+33 more
Domain Name System (DNS) Middleware Identity and Access Management Traceroute Python (Programming Language) CURL Windows Servers Routing Citrix Systems Nslookup Oracle Databases Oracle (Applications) Ping (Networking Utility) Windows PowerShell Reliability Engineering Site Reliability Engineering Practices Cloud Services Ansible Software Engineering SQL Databases Telnet Software Vulnerability Management Data Logging Diagnostic Tools Scripting Load Balancing Software Troubleshooting Firewalls (Computer Science) Information Technology Performance Monitor Oracle Cloud Infrastructure Event Viewer Legacy Systems

Job description

This role combines systems administration, production support, cloud operations, cybersecurity remediation, and site reliability engineering practices. You will work across Windows and Linux systems, applications, cloud infrastructure, networking, and security to resolve technical issues, complete structured deployments, and improve the reliability and maintainability of supported environments. The ideal candidate is a practical problem-solver who is comfortable working from technical runbooks, troubleshooting independently, documenting work thoroughly, and collaborating with infrastructure, application, cybersecurity, networking, cloud, and client-facing teams.

Contributes to the design and architecture of infrastructure and service to ensure reliability and functionality. Responds to infrastructure demands and capacity needs. Collaborates with software development teams to develop of reliable and scalable infrastructures. Contributes to data collection to maintain and optimize operations and reliability. Performs incident response and/or maintenance tasks. Contributes to health and performance reporting. Contributes to identifying opportunities for automation. Communicates relevant details about services and identifies and communicates changes. Provides support for technology and documents incidents. Experiments with new tools and develops working knowledge of site reliability trends., * Administer and support Windows Server and Linux systems in production and non-production environments.

  • Perform post-deployment configuration, application installation, service validation, and environment-readiness checks.
  • Execute structured build and deployment activities using approved runbooks, deployment guides, and change procedures.
  • Run, review, and troubleshoot scripts used for system configuration, deployments, patching, and validation.
  • Investigate operating system, service, application, installation, patching, startup, and connectivity issues.
  • Review logs, Windows Event Viewer, service status, permissions, processes, ports, certificates, and configuration files to identify and resolve problems.
  • Apply operating system, middleware, and application patches, including pre-maintenance checks, post-patch validation, and rollback support.
  • Support vulnerability remediation, system hardening, STIG implementation, and other compliance-driven configuration activities.
  • Troubleshoot network and service-connectivity issues involving DNS, routing, firewalls, load balancers, ports, and certificates.
  • Use diagnostic tools such as ping, traceroute, telnet, netcat, curl, nslookup, and netstat.
  • Support cloud-hosted infrastructure involving compute, storage, networking, identity, access management, and environment provisioning.
  • Participate in incident response, root-cause analysis, and service-restoration activities.
  • Create and maintain technical documentation, runbooks, implementation records, validation results, and escalation notes.
  • Identify opportunities to automate repetitive work, improve operational processes, and reduce recurring incidents.
  • Participate in scheduled maintenance windows, after-hours support, or an on-call rotation as required.

Requirements

  • Bachelor’s degree in Computer Science, Information Technology, Engineering, or a related field, or equivalent practical experience.
  • Three or more years of experience in site reliability engineering, systems administration, infrastructure operations, DevOps, cloud operations, production support, or a related technical field.
  • Hands-on experience administering Windows Server, Linux, or both.
  • Experience installing, configuring, validating, and troubleshooting applications in hosted environments.
  • Experience completing structured build, deployment, configuration, or maintenance activities from technical runbooks.
  • Ability to run, modify, or troubleshoot scripts using one or more of the following:
  • PowerShell
  • Bash
  • Python
  • Ansible
  • Chef
  • Experience applying operating system, middleware, or application patches.
  • Understanding of maintenance windows, change control, rollback planning, and post-maintenance validation.
  • Working knowledge of networking concepts, including DNS, routing, firewalls, load balancers, ports, and certificates.
  • Familiarity with cloud-hosted infrastructure and services.
  • Experience with ticketing, incident-management, or change-management systems.
  • Strong troubleshooting, documentation, and problem-solving skills.
  • Clear written and verbal communication skills.
  • Ability to collaborate effectively across technical and client-facing teams.

Preferred Qualifications

  • Experience with Oracle Cloud Infrastructure.
  • Experience supporting Oracle Health Millennium or Cerner applications and environments.
  • Experience supporting federal clients, government-hosted systems, or regulated environments.
  • Familiarity with STIGs, vulnerability remediation, cybersecurity hardening, and federal compliance workflows.
  • Familiarity with OCI, Scripting, federal compliance, healthcare or Millennium environments
  • Experience with Citrix technologies.
  • Experience supporting legacy systems or complex application architectures.
  • Experience with infrastructure-as-code or configuration-management tools.
  • Familiarity with production monitoring, alerting, centralized logging, and observability practices.
  • Experience with incident response, root-cause analysis, and SRE operational practices.
  • Knowledge of Oracle Database, SQL, middleware, or related Oracle technologies.
  • Relevant certifications in Oracle Cloud Infrastructure, Windows Server, Linux, networking, cybersecurity, or cloud computing.
  • Experience with creating CI/CD pipelines

Only Oracle brings together the data, infrastructure, applications, and expertise to power everything from industry innovations to life-saving care. And with AI embedded across our products and services, we help customers turn that promise into a better future for all. Discover your potential at a company leading the way in AI and cloud solutions that impact billions of lives.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on eeho.fa.us2.oraclecloud.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:06 min

Developer experience and project variety at scale

Alexandra Petri · World Congress 2023

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · World Congress 2025

1:45 min

The ubiquity of the curl library in connected devices

Daniel Stenberg · Coffee With Developers

8:22 min

Simulating a Linux terminal and running Spring Boot

Jakov Semenski · LIVE

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

3:55 min

Demonstrating .NET installation on Debian and Azure Linux

Silvano Coriani Silvano Coriani · Europe 2026 Virtual

Videos

See all

Related articles

See all