Site Reliability Engineer
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+11 more
Requirements
- Requires BS degree and 8 - 12 years of prior relevant experience or master’s with 6 - 10 years of prior relevant experience. Additional years of experience will be considered in lieu of degree.\n
- Proven experience as a Site Reliability Engineer or similar role.\n
- Strong knowledge of cloud platforms (e.g., AWS, Azure, Google Cloud).\n
- Proficiency in scripting and coding (e.g., Python, Bash, Go).\n
- Experience with configuration management tools (e.g., Ansible, Terraform).\n
- Familiarity with containerization and orchestration (e.g., Rancher/Harvester).\n
- Excellent problem-solving and analytical skills.\n
- Strong communication and collaboration abilities.\n
- Serve as a liaison between site operators and consortium DI team members.\n
- Ability to work in an on-call rotation.\n
- Familiar with Cisco/Juniper and Linux CLI and Logs\n
- Knowledge of IdAM and VSAN/Cloud Storage\n, * Experience with CI/CD pipelines and tools (e.g., Jenkins, GitLab CI/CD).\n
- Knowledge of networking and security best practice.\n
- Familiarity with database management and optimization.\n
- Experience with HAIPE devices/KG encryptors and security associations\n
- Experience with OSI Layers 1 - 4 troubleshooting methodologies\n
Benefits & conditions
Leidos has an opening for a highly qualified \nTS/SCI cleared Site Reliability Engineer at \nHickam Air Force Base, Hawaii for the \nDecision Advantage Business Area in Defense sector. This is an exciting opportunity to bring your experience to support all-domain large-scale weapon systems, Information Technology Systems, and Command and Control Systems to realize the Department of Defense Joint All-Domain Command and Control (JADC2). In this role you will support the Advanced Battle Management System (ABMS) Digital Infrastructure (DI) Processing Node (PN) team to design and implement solutions that can be delivered at speed, scale, and with the necessary security to deliver operational advantages to the joint warfighter. ABMS is a top modernization priority for the Department of the Air Force and will be the backbone of a network-centric approach to battle management in partnership with all the services across JADC2. This position will work closely with Program Managers, domain engineers, and Government counterparts across Government and Industry partners. \n \n \nPrimary responsibilities:\n \n \n
- Monitor and maintain system health.\n
- Automate repetitive tasks to improve efficiency and reliability.\n
- Build and maintain effective monitoring and alerting systems.\n
- Collaborate with development teams to ensure smooth deployment of new features.\n
- Troubleshoot and resolve incidents to minimize downtime.\n
- Conduct post-incident reviews and implement improvements.\n
- Optimize system performance and scalability.\n
- Document processes and procedures for future reference.\n
-
Must be able to perform the following (but not limited to) data center tasks:\n, * Required Certifications: \n \n
- DOD 8140 IAT II (Intermediate)\n
- GIAC Security Essentials Certification;\n
- GICSP: Global Industrial Cyber Security Professional: SSCP: Systems Security Certified Practitioner\n \n
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
Highest Paying Tech Companies for Developers
Is Software Engineering Over-Saturated?
Fully Remote Software Engineer Jobs
Find a Developer Job: 12 Best Job Sites For Developers