> Markdown version of [/jobs/ext/611466-site-reliability-engineer-sre](https://www.wearedevelopers.com/jobs/ext/611466-site-reliability-engineer-sre). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer (SRE) - **Company:** CATHEXIS PARTNERS, LLC - **Location:** Tysons, United States - **Experience:** Experienced - **Salary:** $100,000.0 - $160,000.0 - **Contract:** Permanent contract - **Skills:** Microsoft Access, Java (Programming Language), JavaScript (Programming Language), Amazon Web Services, Amazon S3, Apache HTTP Server, Computer Vision, Build Automation, Microsoft Azure, C++ (Programming Language), Cloud Computing, Cloud Engineering, Linux, Distributed Data Store, Hadoop Distributed File System, Python (Programming Language), PostgreSQL, Natural Language Processing, Network File Systems, Object-Oriented Software Development, Open Source Technology, Scrum Methodology, Reliability Engineering, Cloud Services, Ansible, Ruby, Mesos, Software Deployment, Software Engineering, Reinforcement Learning, Ceph (Software), Google Cloud, Cloud Platform System, Apache Yarn, Deep Learning, Kubernetes Helm Charts, Infrastructure as Code (IaC), Cloudformation, Containerization, Kubernetes, Infrastructure Automation Frameworks, Information Technology, Data Analytics, Terraform, Docker, Jenkins, Vulnerability Analysis, Microservices - **Published:** June 24, 2026 - **Apply:** https://www.dice.com/job-detail/38e9f43e-7723-4515-8c8f-4039e4427a43 ## About the Role * Active TOP SECRET clearance or higher is required * Bachelor's degree in Computer Science or related field * A minimum of two (2) years of experience working with on-premise and off-premise cloud environments * Experience with AWS and/or Azure * Hands-on experience with a range of open-source technologies, such as Linux, Docker, Kubernetes, K8s, Terraform, Helm, PostgreSQL, or similar technologies * Ability to program (structured and OOP) using one or more high-level languages, such as Python, Java, C/C++, Ruby, and JavaScript * Experience with distributed storage technologies such as NFS, HDFS, Ceph, and Amazon S3, as well as dynamic resource management frameworks (Apache Mesos, Kubernetes, Yarn) * Proactive approach to identifying problems, performance bottlenecks, and areas for improvement * Ability to lead and work independently in an Agile/Scrum environment * Real passion for developing team-oriented solutions to complex engineering problems * Thrive in an autonomous, empowering and exciting environment * Great verbal and written communication skills to collaborate multi-functionally and improve scalability * Interest in committing to a fun, friendly, expansive, and intellectually stimulating environment Desired Skills * Hands-on experience deploying and operating applications using IaaS and PaaS on major cloud providers, such as Amazon AWS, Microsoft Azure, or Google Cloud Services * Experience with deep learning, natural language processing, computer vision, or reinforcement learning * Conveys highly technical concepts and information in written form to technical and non-technical audiences * The ability to work on multiple concurrent projects is essential. Strong self-motivation and the ability to work with minimal supervision * Must be a team-oriented individual, energetic, result & delivery oriented, with a keen interest on quality and the ability to meet deadlines ## Description Team CATHEXIS elevates the government contracting experience through rapid response, deep skill, and thoughtful problem-solving and communication. Our core capabilities are our top-tier program and project management, data analytics, and audit services, the backbone of which is our integrated approach to operational excellence. You worked hard to get to where you are. You strive to make every day better than the day before. So do we. Team CATHEXIS operates with an all-in mindset. We are working together to create a company that supports our shared values and individual goals. Our values are centered around leading with integrity, owning the outcome, growing together, and moving with purpose in everything we do for our employees, customers, partners, and communities. We believe success is best when we listen and lead with empathy; model high standards of ethics to provide a rewarding candidate experience; work hard, have fun, and appreciate the strengths we all bring to the team; and empower our employees to create innovative and trusted results. We are looking for a dynamic Site Reliability Engineer (SRE) with a Top Secret clearance to join our team! The Site Reliability Engineer (SRE) will manage, monitor, and optimize clusters on Kubernetes. Together, we're accelerating our clients' digital transformation through the building and deployment of data-driven, scalable AI solutions. The ideal candidate will have a deep understanding of Kubernetes, Cloud Infrastructure, and Infrastructure as Code (IaC) practices. You will be responsible for ensuring the reliability and scalability of our clients' Kubernetes clusters and Cloud Infrastructure. We are proactively building a pipeline of qualified candidates for future opportunities that may arise. If your background aligns with our anticipated needs, a member of our Talent Acquisition team may reach out should a role become available or to proactively screen you for the role., * Monitor and Manage Kubernetes Clusters: Ensure the stability, health, and scalability of Kubernetes Clusters, deploying applications and services on Kubernetes * Kubernetes Management: Deploy, monitor, and scale applications on Kubernetes clusters. Maintain Helm charts, manage services, and ensure resource allocation for optimal cluster performance * Containerization & Deployment: Design and maintain Docker-based microservices architecture, ensuring consistent and reproducible deployments across staging, QA, and production environments * Cloud Infrastructure Management: Work with leading Cloud Platforms (AWS, Azure and/or Google Cloud Platform) to set up, configure, and manage infrastructure resources using Infrastructure as Code (Terraform, CloudFormation, etc.) * Monitoring & Incident Response: Set up monitoring solutions, define alerts, an manage the incident response process for any issues related to Jenkins or Kubernetes clusters * Automate Infrastructure Processes: Build automation tools for scaling, monitoring, and maintaining infrastructure using modern tools like Terraform, Ansible, Linux, or equivalent * Collaborate Across Teams: Work closely with development, services, and operations teams to ensure a seamless integration between application development, deployment, and infrastructure * Security & Compliance: Ensure all systems follow best practices in terms of security and compliance with relevant regulations. This includes role-based access, encryption, and automated vulnerability scanning ## Related Videos - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [Coffee with Developers: David Heinemeier Hansson](https://www.wearedevelopers.com/videos/875-coffee-with-developers-david-heinemeier-hansson) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [Fireside Chat with Werner Vogels, VP & CTO, Amazon.com & Daniel Gebler, CTO at Picnic](https://www.wearedevelopers.com/videos/1405-fireside-chat-with-werner-vogels-vp-cto-amazon-com-daniel-gebler-cto-at-picnic) - [Docker exec without Docker](https://www.wearedevelopers.com/videos/1094-docker-exec-without-docker) ## Related Articles - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [What Are The Top Skills Required For Azure Developers?](https://www.wearedevelopers.com/magazine/77-what-are-the-top-skills-required-for-azure-developers) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Why Upskilling And Reskilling is Important For Developers](https://www.wearedevelopers.com/magazine/428-why-upskilling-and-reskilling-is-important-for-developers) - [Résumé-Driven Development: How IT trends affect the job market for software developers](https://www.wearedevelopers.com/magazine/59-resume-driven-development-how-it-trends-affect-the-job-market-for-software-developers)