Director - Splunk Platform Engineering & SRE

Capgemini
New York, NY, United States
8 days ago

Role details

Contract type
Temporary contract
Employment type
Full-time (> 32 hours)
Compensation
$144,227.0 - $225,368.0
Working hours
Regular working hours
Job source

Tech stack

Java (Programming Language) Artificial Intelligence Data Analysis User Authentication Unix Cloud Computing Configuration Management Cyber Security Computer Programming Software Debugging Linux DevOps
+29 more
Disaster Recovery Distributed Systems Monitoring of Systems Python (Programming Language) Network Layer Linux System Administration Machine Learning Networking Basics Packet Analyzer Performance Tuning Role-Based Access Control Reliability Engineering Ansible Prometheus Security Information and Event Management Syslog Systems Integration TCP/IP Scripting Data Ingestion Git Containerization Kubernetes Infrastructure Automation Frameworks Information Technology Low Latency Data Management Splunk Golang

Job description

New York, NY, United States (On-site) Contract (11 months 28 days) Published 16 hours ago Monitoring & Observability linux root cause analysis Linux administration networking skills devops SIEM platforms kubernetes AI/ML python programming

  • We’re seeking a team member for the role of Director, Splunk Platform Engineering & SRE (Individual Contributor) to join our Cybersecurity Engineering Tools & Platforms team. This role is in New York, NY
  • This is a high-impact, deeply technical leadership role designed for a top-tier engineer, not a people manager. The Director title reflects the depth of technical expertise, ownership, and influence required, not team size.
  • You will take ownership of a large-scale, mission-critical Splunk platform at the center of enterprise observability and cybersecurity. This role requires someone who can go deep into the stack: OS, network, ingestion pipelines, distributed systems, and resolve issues at their root, regardless of complexity.
  • If you are the engineer that others call when systems fail in unpredictable ways, and you enjoy solving those problems, this role is built for you.

In this role, you’ll make an impact in the following ways:

  • Own end-to-end engineering and operational accountability for the enterprise Splunk platform (SIEM), including architecture, capacity planning, ingestion, integrations, and lifecycle management
  • Act as the highest technical escalation point, driving resolution of critical incidents across application, platforms, and infrastructure layers

Troubleshoot and resolve deep, low-level technical issues, including:

  • Linux/Unix OS internals (CPU, memory, I/O, process behavior)
  • Network behavior, packet flow, and latency bottlenecks
  • Distributed system failures and data ingestion breakdowns
  • Drive platform reliability, capacity, observability, and performance engineering, using modern monitoring stacks (Prometheus, Moog)

Architect and scale high-throughput ingestion pipelines, integrating:

  • Syslog and event ingestion frameworks
  • Kubernetes / containerized platforms
  • Cloud and enterprise systems
  • Own authentication, RBAC, and access control models, ensuring strong governance and compliance
  • Design and implement automation and configuration management frameworks (Git, Ansible) to reduce operational toil
  • Lead incident response, root cause analysis, and systemic fixes, embedding SRE principles (SLAs, SLOs, error budgets)
  • Drive platform upgrades, resilience strategies, and disaster recovery readiness
  • Evaluate and onboard emerging technologies, including AI/ML-driven analytics and contextual data platforms
  • Create bespoke solutions for unsolved problems using languages like python, java or golang.
  • Influence engineering direction across teams through technical leadership and expertise as an individual contributor
  • Mentor and elevate engineers through hands-on guidance and technical depth

Requirements

  • Bachelor’s degree in computer science or a related discipline, or equivalent work experience required, advanced degree preferred.
  • 12+ years of experience in information security or related technology experience required, experience in the securities or financial services industry is a plus.
  • Strong foundation in Site Reliability Engineering (SRE) and distributed systems
  • Proven ability to debug and resolve complex issues across the full stack, from application to OS and network layers
  • Expert knowledge of Linux/Unix systems, including performance tuning and low-level troubleshooting
  • Strong understanding of networking fundamentals (TCP/IP, packet analysis, syslog pipelines, latency debugging)
  • Experience building and operating high-volume data ingestion and processing systems
  • Proficiency in Splunk SPL, and data analysis
  • Strong programming/scripting skills (e.g., Python, Go, Java, or similar)
  • Hands-on experience with DevOps and configuration management tools (Ansible, Git, etc.)
  • Experience with Kubernetes and containerized environments
  • Deep understanding of security models, RBAC, and enterprise controls
  • Ability to operate independently in high-pressure situations and take full ownership of outcomes
  • A mindset focused on automation, scalability, and eliminating operational friction
  • Technical in depth and hands-on A.I. literacy as well as knowledge of MCP design

Benefits & conditions

Success Profile:

  • Becomes the go-to technical authority for one of the firm’s most critical platforms
  • Resolves high-impact, complex incidents quickly and effectively
  • Improves platform performance, scalability, and resilience at scale
  • Reduces manual operational burden through engineering and automation
  • Strengthens security posture and access governance
  • Raises the technical bar across the organization through expertise and influence

The pay range that the employer in good faith reasonably expects to pay for this position is $69.34/hour - $108.35/hour. Our benefits include medical, dental, vision and retirement benefits. Applications will be accepted on an ongoing basis.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

2:03 min

Microsoft integrating native Unix coreutils into Windows environments

Chris Heilmann +2 · LIVE

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · WWC 2025

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

56 sec

Favorite git commands and the importance of patch commits

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

4:01 min

Finding personal fulfillment in the cybersecurity industry

LIVE

Videos

See all

Related articles

See all