Director, Deployment Engineering - Systems Engineering

Nscale
Barcelona, Spain
2 days ago
Apply on www.buscojobs.com.es
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience required
10 years minimum
Working hours
Regular working hours

Tech stack

Artificial Intelligence Systems Engineering Unix Configuration Management Data Centers Data Synchronization Linux Programming Tools Distributed Systems Fault Tolerance Monitoring of Systems Performance Tuning
+16 more
Ansible Systems Architecture Virtualization Technology AI Infrastructure Iq/oq/pq Scripting Graphics Processing Unit (GPU) Reliability of Systems Containerization Kubernetes Information Technology Operational Systems Data Management Puppet Terraform Docker

Job description

AMERThe roleNscale is looking for a Director of Deployment Engineering to lead the systems engineering function responsible for deploying, validating, and operating the core compute platforms that underpin our AI infrastructure.You will lead a high-performing engineering team across systems, compute operations, and deployment validation. This role combines technical leadership with operational execution: defining the standards, tooling, and processes that ensure infrastructure is reliable, scalable, production-ready, and delivered at pace.What you’ll doLead and scale the systems engineering team responsible for core platform systems, deployment readiness, and compute operations.Define and deliver multi-quarter systems engineering initiatives that improve deployment velocity, platform reliability, validation quality, and operational performance.Establish systems validation standards for servers, GPUs, networking, storage, and supporting infrastructure before production acceptance.Lead GPU burn-in and validation testing at scale, including thermal, power, stress, and performance qualification.Partner with infrastructure, network engineering, deployment, product, and operations leaders to balance long-term platform architecture with immediate business needs.Turn ambiguous, high-impact technical challenges into clear plans, priorities, milestones, and accountable execution.Drive alignment across interdependent teams working on compute platforms, internal infrastructure, developer tooling, and operational systems.Raise the bar for engineering quality, automation, observability, documentation, and operational excellence.Build scalable approaches for systems monitoring, telemetry, incident learning, and continuous platform improvement.Travel up to 50% to support deployments, site readiness, vendor collaboration, and operational execution.What you’ll bringA bachelor’s degree in Computer Science, Engineering, or a related technical field.10+ years of systems engineering, compute operations, infrastructure, or engineering-management experience.Experience in a large cloud provider, hyperscale data center, or similarly complex infrastructure environment.Experience leading systems or compute operations teams in a high-availability production environment.Strong knowledge of Linux/Unix systems administration, OS-level tuning, server architecture, and GPU hardware.Experience with virtualization, containerization, and distributed systems, including Kubernetes and Docker.Experience with infrastructure-as-code and configuration-management tools such as Ansible, Terraform, Puppet, or Chef.Strong scripting, automation, and data center systems-design experience.Familiarity with system architecture, data synchronization, fault tolerance, state management, and distributed-system reliability.Experience with enterprise storage, networking, compute, monitoring, observability, or telemetry platforms.Excellent judgment, organizational skills, and written and verbal communication.The ability to influence technical direction, engineering priorities, and cross-functional decisions.What success looks likeNscale’s compute platforms are consistently deployed, validated, and accepted into production at a high standard.Systems engineering has clear technical standards, automation, and operational ownership across the deployment lifecycle.Platform reliability, validation quality, and delivery velocity improve measurably over time.Engineering, operations, and deployment teams make faster decisions with clearer data and accountability.The range below reflects the base salary for the position. Actual compensation may vary based on job-related factors such as skill set, experience, education, and location. In addition to base salary, this role may be eligible for bonus, equity, and/or commission programs. Nscale may offer a competitive benefits package including medical, dental, vision, flexible paid time off, parental leave, and retirement plan participation.Salary Range$240,000 - $353,000 USD#J-*****-Ljbffr

Requirements

A bachelor’s degree in Computer Science, Engineering, or a related technical field. 10+ years of systems engineering, compute operations, infrastructure, or engineering-management experience. Experience in a large cloud provider, hyperscale data center, or similarly complex infrastructure environment. Experience leading systems or compute operations teams in a high-availability production environment. Strong knowledge of Linux/Unix systems administration, OS-level tuning, server architecture, and GPU hardware. Experience with virtualization, containerization, and distributed systems, including Kubernetes and Docker. Experience with infrastructure-as-code and configuration-management tools such as Ansible, Terraform, Puppet, or Chef. Strong scripting, automation, and data center systems-design experience. Familiarity with system architecture, data synchronization, fault tolerance, state management, and distributed-system reliability. Experience with enterprise storage, networking, compute, monitoring, observability, or telemetry platforms. Excellent judgment, organizational skills, and written and verbal communication. The ability to influence technical direction, engineering priorities, and cross-functional decisions.

Benefits & conditions

Nscale’s compute platforms are consistently deployed, validated, and accepted into production at a high standard. Systems engineering has clear technical standards, automation, and operational ownership across the deployment lifecycle. Platform reliability, validation quality, and delivery velocity improve measurably over time. Engineering, operations, and deployment teams make faster decisions with clearer data and accountability. The range below reflects the base salary for the position. Actual compensation may vary based on job-related factors such as skill set, experience, education, and location. In addition to base salary, this role may be eligible for bonus, equity, and/or commission programs. Nscale may offer a competitive benefits package including medical, dental, vision, flexible paid time off, parental leave, and retirement plan participation. Salary Range $240,000 - $353,000 USD #J-*****-Ljbffr

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.buscojobs.com.es
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:03 min

Microsoft integrating native Unix coreutils into Windows environments

Chris Heilmann Chris Heilmann +2 · LIVE

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · World Congress 2025

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · World Congress 2025

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

2:04 min

Defining timestamps and the international standard format

Denny Biasiolli Denny Biasiolli · Europe 2026 Virtual

4:41 min

Scale and diversity of software development teams

Bastian Heilemann Bastian Heilemann +1 · World Congress 2025

Videos

See all

Related articles

See all