> Markdown version of [/jobs/ext/2946172-sr-manager-engineering-data-infrastructure-mlops-hybrid](https://www.wearedevelopers.com/jobs/ext/2946172-sr-manager-engineering-data-infrastructure-mlops-hybrid). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Sr. Manager, Engineering - Data Infrastructure & MLOps (Hybrid) - **Company:** CrowdStrike - **Location:** Austin, TX, United States (Remote available) - **Experience:** Expert - **Salary:** $160,000.0 - $250,000.0 - **Contract:** Permanent contract - **Skills:** Java (Programming Language), Artificial Intelligence, Apache HTTP Server, Big Data, Data Centers, Data Infrastructure, Data Recovery, Data Systems, Memory Management, Amazon DynamoDB, Apache Hadoop, Apache Hive, Java Virtual Machine (JVM), PostgreSQL, MySQL, Online Analytical Processing, NoSQL, Performance Tuning, Ansible, Standard Sql, Scala (Programming Language), Software Engineering, Data Streaming, Enterprise Software Applications, Real Time Systems, Apache Spark, HybridCloud, Kotlin, Falcon Platform, Data Lakes, Kubernetes, Information Technology, Druid, Apache Flink, Deployment Automation, Cassandra, Bare Metal, Data Analytics, Apache Kafka, Data Management, Machine Learning Operations, Presto, Puppet, Terraform - **Published:** September 16, 2026 - **Apply:** https://www.jofdav.com/jobs/59724862-sr-manager-engineering-data-infrastructure-mlops-hybrid ## About the Role * MS in Computer Science or related field * 10+ years of software engineering experience in building and/or managing large scale data platforms * 5+ years of hands-on management experience leading distributed engineering teams * Experience building and supporting real time or NRT data systems of ultra high scale (MINIMUM: 100s of TB/day in either a current or past role) * Significant expertise in performance-tuning or developing internals (source code) within Spark, Flink, Iceberg or Pinot or an equivalent structured streaming, Time Series, OLAP or Open Table real time system. * Production experience building either Spark- or Flink-based self-service data platforms, or equivalent (i.e. with Ray, or building spark- or Flink-like frameworks themselves in Scala, Akk, etc.) * Strong familiarity with (and ample hands-on experience tuning & optimizing) at least one applicable technology in the Apache Hadoop ecosystem: Spark, Kafka, Hive/Iceberg/Delta Lake, Presto/Trino, Pinot, Druid, etc. * Ample experience coding in Java, Scala, Kotlin or another JVM language (bonus points for experience tuning the language, i.e. garbage collection, memory management…) * Production experience with relational SQL and NoSQL databases, including Postgres/MySQL, Cassandra/DynamoDB, etc. * Intimate knowledge of deployment tools like Terraform and configuration automation tools like Chef/Puppet/Ansible * Demonstrated track record of leadership and building a strong core engineering team * Building and managing cloud and on-premise Hadoop and Spark infrastructure which could run on bare metal or Kubernetes clusters * Strong cross-team collaboration and interpersonal skills working with various roles including engineering, product management, support and sales engineering * Demonstrated ability to attract and hire talent and grow the team rapidly * Experience working with remote teams and individuals while ensuring agility and code velocity * Ability to communicate and articulate crisply at all levels from executive staff to engineers * Broad general knowledge of the high-technology industry gained in larger enterprise software environments enhanced by ongoing awareness of R&D practices / technology advances * Proven experience utilizing AI technologies to enhance decision-making, streamline workflows and processes, improve efficiency and drive business outcomes. Bonus Points: * Management experience with any of the following, the more the better * Developing compute and storage layers for data platforms at high scale (batch or near Real-time) * Experience with hybrid cloud environments (cloud and data center) * Exposure to container orchestration technologies, like Kubernetes * Experience with remote work, particularly managing distributed teams across multiple timezones #LI-MP2 ## Description CrowdStrike is seeking a Sr Engineering Manager (L8) to lead a host of initiatives spanning 2-3 teams with coverage in MLOps & broader data infrastructure, as well as data analytics platform & data recovery initiatives. Your teams would own data plane infrastructure that runs a multitude of Spark-, Flink- and other apache-based applications at incredibly high scale (over 3 Trillion events/day) both in our cloud environments as well as data centers. You own the setup, deployment, scaling and management of all data plane products that are used to ingest, process and enrich event stream data that are part of the Falcon platform. In addition, you will continue to build out & scale teams capable of supporting the myriads of services running within the our data infrastructure. You will work closely with engineers and product to understand requirements and cross functionally with peers and senior leadership to deliver innovative products that can scale with increased demand. CrowdStrike is a remote-first organization and this position requires the ability to hold meetings with team members across US and EU time zones. This is a Hybrid role requiring an 2 days per week in one of our office Locations (Austin, TX or Mid-town Manhattan, NY) What You'll Do: * Lead multiple teams in the design, scaling and maintenance of data infrastructure services operating at 100s of PBs scale * Train, mentor and continue to hire top-tier data platform engineering talent ## Related Videos - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [Leveraging Real time data in FSIs](https://www.wearedevelopers.com/videos/806-leveraging-real-time-data-in-fsis) - [MySQL Protocol Features You Should Be Aware Of](https://www.wearedevelopers.com/videos/100267-mysql-protocol-features-you-should-be-aware-of) - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [NoSQL Data Modeling for Front-end Developers](https://www.wearedevelopers.com/videos/297-nosql-data-modeling-for-front-end-developers) - [Coding for Good: Achieving social change with an app](https://www.wearedevelopers.com/videos/1645-coding-for-good-achieving-social-change-with-an-app) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers)