> Markdown version of [/jobs/ext/2184120-staff-site-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/2184120-staff-site-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Staff Site Reliability Engineer - **Company:** NINJATRADER, LLC - **Location:** Chicago, IL, United States (Remote available) - **Experience:** Expert - **Salary:** $160,000.0 - $210,000.0 - **Contract:** Permanent contract - **Skills:** Amazon Web Services, Microsoft Azure, Bash Shell, Continuous Integration, DevOps, Github, Monitoring of Systems, Identity and Access Management, Python (Programming Language), Open Source Technology, PCI Data Security Standards, Systems Development Life Cycle, Reliability Engineering, Ansible, Prometheus, Datadog, Google Cloud, Cloud Platform System, Grafana, Kubernetes, Infrastructure Automation Frameworks, Deployment Automation, Build Tools, Terraform, Docker - **Published:** August 22, 2026 - **Apply:** https://www.dice.com/job-detail/cb59c304-cfe8-4279-9578-081b98de4c61 ## About the Role * 8+ years of experience in DevOps, Site Reliability Engineering, or Platform Engineering roles * Expertise with Kubernetes, Docker, and container orchestration * Hands-on experience with CI/CD tools (GitHub Actions or equivalent) * Proficiency in programming languages (e.g., Python, Bash, or Go) and automation tools such as Ansible, Terraform, or Helm * Hands-on experience with AWS, Google Cloud Platform, or Azure, including in-depth knowledge of networking, security, and identity management in cloud environments * Knowledge of monitoring and observability tools such as Prometheus, Grafana, Datadog, or similar * Strong collaboration, communication, and leadership skills, with the ability to influence technical decisions across teams and mentor junior engineers Bonus points for: * Trading industry experience * Contributions to open-source projects ## Description * Serve as the technical lead for the SRE function, setting technical direction and mentoring engineers across reliability initiatives * Analyze, troubleshoot, and remediate production issues with a systematic problem-solving approach to keep revenue-generating systems running * Participate in a weekly 12x7 on-call rotation, including weekend deployments and checkouts before markets open on Sundays * Perform initial root cause analysis and remediation of production incidents * Build tools that automate repetitive tasks, deployments, and incident responses with the goal of minimal human involvement * Design reliable monitoring and alerting systems with Product and QA teams, establishing and tracking SLIs/SLOs across web, mobile, desktop, and trading platforms * Leverage Infrastructure as Code tools like Terraform to automate provisioning, scaling, and management across all platforms * Collaborate with Product Engineering, Operations, and other cross-functional teams to deliver features on time while meeting scalability, security, and performance requirements * Implement security and compliance best practices (e.g., SOC 2, PCI DSS) throughout the software delivery lifecycle for web, mobile, and desktop deployments ## Related Videos - [Inside Bitpanda's Tech Stack: Scaling a European Fintech Leader - Markus Dorner](https://www.wearedevelopers.com/videos/1979-inside-bitpanda-s-tech-stack-scaling-a-european-fintech-leader-markus-dorner) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [Innovating Developer Tools with AI: Insights from GitHub Next](https://www.wearedevelopers.com/videos/1268-innovating-developer-tools-with-ai-insights-from-github-next) - [My journey into DevOps world - How it all started!](https://www.wearedevelopers.com/videos/545-my-journey-into-devops-world-how-it-all-started) - [#90DaysOfDevOps - The DevOps Learning Journey](https://www.wearedevelopers.com/videos/548-90daysofdevops-the-devops-learning-journey) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [The Best Software Developer Blogs to Read](https://www.wearedevelopers.com/magazine/156-the-best-software-developer-blogs-to-read)