> Markdown version of [/jobs/ext/2124453-staff-infrastructure-engineer](https://www.wearedevelopers.com/jobs/ext/2124453-staff-infrastructure-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Staff Infrastructure Engineer - **Company:** · Sentinelone - **Location:** New York, NY, United States (Remote available) - **Experience:** Expert - **Salary:** $156,000.0 - $215,000.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Data Analysis, Microsoft Azure, Cloud Computing, Databases, Continuous Delivery, Continuous Integration, Data Infrastructure, Distributed Data Store, Distributed Systems, Github, Home Automation, Internet Security, Python (Programming Language), Octopus Deploy, Open Source Technology, Platform as a Service (PAAS), Performance Tuning, Redis, Pulumi, Scripting, Multi-Cloud, SC Clearance, Kubernetes, Deployment Automation, Apache Kafka, Data Management, Terraform, Multiplatform, SentinelOne Expertise, Jenkins, Golang - **Published:** August 19, 2026 - **Apply:** https://www.careerbuilder.com/job-details/staff-infrastructure-engineer-data-streaming-new-york-ny--f717673d-bf5d-459a-a5ab-a4093eb5109c ## About the Role * 8 or more years of experience in infrastructure/platform engineering, with a proven track record of operating stateful distributed systems at scale. * Deep hands-on experience with self hosted Kafka and Redis running in Kubernetes, including performance tuning, scaling, partitioning, persistence, and operator-based lifecycle management. * Strong understanding of Kubernetes internals and best practices for managing both stateless and stateful workloads in production environments. * Experience providing Database- or Messaging-as-a-Service (DBaaS/PaaS) for internal development teams or external customers. * Exposure to multi-cloud environments with strong expertise in at least one major provider: AWS, GCP, or Azure. * Experience with Infrastructure as Code and GitOps practices (Terraform, ArgoCD, Pulumi). * Familiarity with advanced deployment strategies (blue-green, canary, rolling). * Strong scripting or development skills (e.g., Python, Go, or similar). * Solid understanding of CI/CD pipelines and workflow automation (GitHub Actions, Argo Workflows, etc.)., Amazon Web Services (AWS), Apache Kafka, Artificial Intelligence (AI), Automation, Best Practices, Cancer, Cellular Telephone, Cloud Computing, Coaching, Continuous Deployment/Delivery, Continuous Integration, Cost Control, Data Analysis, Data Management, Distributed Computing, Editing, Employee Assistance Plan, Endpoint Security, Flexible Spending Accounts, GCP (Good Clinical Practices), GitHub, Go Programming Language (Golang), Health Insurance, High Throughput, Home Automation, Insurance, International Business, Internet Security, Jenkins, Legal, Microsoft Windows Azure, Multiplatform/Cross-Platform, Open Source, Pager, Performance Tuning/Optimization, Platform as a Service (PaaS), Problem Solving Skills, Production Systems, Python Programming/Scripting Language, Redis, Reimbursement, Scripting (Scripting Languages), Secret Clearance, Standards Development, Stock Purchase Plans, Willing to Travel ## Description * Lead the design and operation of distributed data services, including Kafka and Redis (self hosted), running at massive scale across Kubernetes clusters and multi-cloud environments. * Unlock complete cloud portability for SentinelOne's services by building a highly automated, self-service infrastructure that can run seamlessly across AWS, GCP, and air-gapped on-prem environments. * Manage data infrastructure supporting 5 or more PB/day ingestion, ensuring low-latency, high-throughput, and cost-effective operation at global scale. * Consolidate and optimize multi-tenant Kafka clusters to reduce cost, improve resilience, and streamline operations. * Drive Redis and Kafka lifecycle automation using GitOps principles (ArgoCD, Terraform), reducing operational toil and minimizing pager fatigue. * Define and implement standards for observability, HA, backup, and DR of stateful workloads in Kubernetes. * Partner with FinOps and engineering stakeholders to continuously optimize performance, cost, and operational overhead across data platform components. * Own the end-to-end platform experience for mission-critical open-source systems, serving hundreds of product teams. * Collaborate closely with teams in Europe and India (US Eastern Time Zone preferred due to these collaboration requirements), working with orchestration tools such as Kubernetes (EKS, GKE), Jenkins, GitHub Actions, ArgoCD, and Terraform. ## Related Videos - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [Why segmenting your infrastructure into tiers makes your infrastructure design better](https://www.wearedevelopers.com/videos/1960-why-segmenting-your-infrastructure-into-tiers-makes-your-infrastructure-design-better) - [Go with the Flow: Stop the Leaks Before Your Memory's a Waterfall!](https://www.wearedevelopers.com/videos/100073-go-with-the-flow-stop-the-leaks-before-your-memory-s-a-waterfall) - [Reducing LLM Calls with Vector Search Patterns - Raphael De Lio (Redis)](https://www.wearedevelopers.com/videos/1714-reducing-llm-calls-with-vector-search-patterns-raphael-de-lio-redis) - [Unleashing Potential Across Teams: The Power of Infrastructure as Code](https://www.wearedevelopers.com/videos/930-unleashing-potential-across-teams-the-power-of-infrastructure-as-code) - [How I saved 200K/yr in direct costs writing 0 code lines in K8s](https://www.wearedevelopers.com/videos/1055-how-i-saved-200k-yr-in-direct-costs-writing-0-code-lines-in-k8s) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Why Upskilling And Reskilling is Important For Developers](https://www.wearedevelopers.com/magazine/428-why-upskilling-and-reskilling-is-important-for-developers) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers)