> Markdown version of [/jobs/ext/3080486-site-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/3080486-site-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer - **Company:** Space Exploration Technologies Corp. - **Location:** Hawthorne, United States - **Salary:** $145,000.0 - **Contract:** Permanent contract - **Skills:** Information Systems, Databases, Linux, DevOps, Python (Programming Language), Machine Learning, Reliability Engineering, Software Engineering, Scripting, Cloud Platform System, Kubernetes, Information Technology - **Published:** September 25, 2026 - **Apply:** https://startup.jobs/site-reliability-engineer-kubernetes-platform-top-secret-clearance-spacex-10191658 ## About the Role * Bachelor's degree in computer science, information systems, or an engineering discipline; OR 2+ years of professional experience in software, DevOps, or site reliability engineering in lieu of a degree * 1+ year of experience with Kubernetes * 1+ year of experience with Linux operating systems * Experience in Bash, Python, and/or other scripting languages * Experience building, maintaining, and scaling on-premises and/or cloud systems * Active Top Secret, Top Secret SCI, or DOE Level Q clearance PREFERRED SKILLS AND EXPERIENCE: * Experience hosting and pushing the state of the art in inferential model benchmarks * Experience with systems administration, site reliability engineering, or DevOps engineering * Experience with Python and Python-based development frameworks * Experience with virtualization and hypervisor technologies * Experience with automatically managing dozens or hundreds of servers * Knowledge of performance bottlenecks and performance improvement techniques * Excellent communications skills with the ability to communicate with customers, peers, management etc. in both formal and informal situations * Ability to quickly learn new tools and frameworks., * To conform to U.S. Government export regulations, applicant must be a (i) U.S. citizen or national, (ii) U.S. lawful, permanent resident (aka green card holder), (iii) Refugee under 8 U.S.C. § 1157, or (iv) Asylee under 8 U.S.C. § 1158, or be eligible to obtain the required authorizations from the U.S. Department of State. Learn more about the ITAR here. ## Description As a member of the Classified IT Systems Engineering team, the Site Reliability Engineer is involved in designing scalable systems capable of supporting a growing volume of data products being generated in mass. We build tools that enable us to work more efficiently, and that help us build software systems that are secure, reliable, and autonomous. Our engineers are responsible for the life cycle of the systems they create, including development, testing, and operational support., * Develop automation to deploy and manage compute resources both on-premises and in the cloud * Build, maintain, and scale on-premises hardware systems designed to host GPU-accelerated machine learning workloads * Deploy and manage core infrastructure such as databases, monitoring and storage * Closely collaborate with software engineers to create highly scalable, operable and maintainable products * Engage in and improve the whole lifecycle of services -- from inception and design, through deployment, operation and refinement ## Related Videos - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [Kubernetes and Microservices with Multi-Model Databases](https://www.wearedevelopers.com/videos/382-kubernetes-and-microservices-with-multi-model-databases) - [From Space to Software: Reliability Lessons 40 Years After Challenger](https://www.wearedevelopers.com/videos/100283-from-space-to-software-reliability-lessons-40-years-after-challenger) - [#90DaysOfDevOps - The DevOps Learning Journey](https://www.wearedevelopers.com/videos/548-90daysofdevops-the-devops-learning-journey) - [Docker exec without Docker](https://www.wearedevelopers.com/videos/1094-docker-exec-without-docker) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Best Countries for Software Engineers](https://www.wearedevelopers.com/magazine/267-best-countries-for-software-engineers) - [Résumé-Driven Development: How IT trends affect the job market for software developers](https://www.wearedevelopers.com/magazine/59-resume-driven-development-how-it-trends-affect-the-job-market-for-software-developers)