Director, Deployment Engineering

Nscale
Girona, Spain
1 day ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience required
10 years minimum
Working hours
Regular working hours

Tech stack

Artificial Intelligence Border Gateway Protocol Big Data Cloud Computing Data Centers Data Synchronization Dynamic Host Configuration Protocol Programming Tools Distributed Systems Domain Name System (DNS) Ethernet Fault Tolerance
+18 more
InfiniBand IPv4 IPv6 Virtual Private Networks (VPN) Multi-protocol Systems Internet Service Provider Network Monitoring Network Protocols Open Shortest Path First (OSPF) TCP/IP Scripting Transport Layer Security Computer Network Technologies Data Center Networking Infrastructure Automation Frameworks Information Technology Data Analytics Data Management

Job description

About NscaleNscale is the GPU cloud engineered for AI.We provide high-performance, scalable infrastructure to AI startups and large enterprise customers, reducing the complexity of developing, deploying, and operating AI workloads.At Nscale, we operate with urgency, ownership, and accountability.We value open communication, practical problem-solving, and engineering excellence as we build the infrastructure powering the future of AI.About the RoleWe are seeking a Director of Network & Compute Engineering to lead and scale the teams responsible for the network and compute platforms underpinning Nscale’s GPU cloud infrastructure.You will define the technical and operational direction for highly available systems spanning data center networking, GPU fabrics, compute infrastructure, platform services, and automation.Working closely with Product Management and engineering leaders, you will translate complex and ambiguous challenges into clear strategies and executable plans.This is a highly visible leadership role requiring deep infrastructure expertise, strong cross-functional judgment, and the ability to balance long-term platform architecture with immediate business and customer needs.This position requires up to 50% travel to Nscale offices, data centers, partner locations, and customer sites.What You’ll Be DoingLead, develop, and scale a high-performing engineering organization responsible for Nscale’s core network and compute platforms.Define and execute multi-quarter initiatives that improve deployment velocity, infrastructure reliability, network validation, platform quality, and operational scalability.Establish the technical strategy for data center networking, GPU fabrics, compute systems, platform services, and supporting automation.Partner with Product Management and engineering leaders to evolve the infrastructure platform as a product, balancing long-term architecture with near-term customer and business requirements.Turn ambiguous, high-impact challenges-including platform scalability, vendor integration, infrastructure abstractions, and deployment consistency-into clear execution plans.Drive alignment across teams working on complex, interdependent systems spanning networking, compute, internal platforms, infrastructure automation, and developer tooling.Establish engineering standards for architecture, validation, observability, change management, reliability, and operational readiness.Improve the consistency and quality of network and compute deployments through automation, testing, telemetry, and data-driven decision-making.Provide technical and operational leadership during critical infrastructure incidents and complex escalations.Partner with hardware manufacturers, network vendors, data center operators, and other strategic partners to deliver infrastructure at scale.Build organizational capability through hiring, coaching, succession planning, and the development of engineering leaders.Raise the bar for engineering quality, operational excellence, accountability, and execution across the organization.Travel up to 50% to support infrastructure deployments, partner engagement, and operational priorities.Minimum QualificationsBachelor’s degree in Computer Science, Engineering, or a related technical field, or equivalent practical experience.10+ years of experience leading network, compute, infrastructure, or platform engineering teams.Experience operating within a large cloud provider, internet service provider, hyperscale data center, or similarly complex infrastructure environment.Experience leading network or compute operations teams responsible for business-critical production environments.Demonstrated experience building and scaling engineering teams and delivering complex, multi-quarter infrastructure initiatives.Strong understanding of large-scale data center, cloud, network, or compute architecture.Ability and willingness to travel up to 50%.Preferred QualificationsExperience designing or operating high-performance GPU networks using InfiniBand or Ethernet-based RoCE.Strong knowledge of networking protocols and technologies including BGP, OSPF, IS-IS, MPLS, TCP/IP, IPv4, IPv6, DNS, DHCP, VPN, and SSL.Experience with overlay networking technologies, including VXLAN and EVPN.Experience with server and GPU hardware architecture, lifecycle management, and systems management.Familiarity with system-level architecture, distributed systems, data synchronization, state management, fault tolerance, and high-availability design.Experience with infrastructure automation, scripting, network validation, and data center design.Experience implementing network monitoring, observability, and telemetry platforms.Broad experience across enterprise storage, networking, compute, and cloud infrastructure.Proven ability to resolve complex technical and organizational challenges using sound judgment and creative problem-solving.Experience influencing product roadmaps, technical priorities, and platform investment decisions.Excellent organizational, written, and verbal communication skills.Ability to influence senior stakeholders and create alignment across engineering, product, operations, vendors, and customers.The range below reflects the base salary for the position.Actual compensation may vary based on job-related factors such as skill set, experience, education, and location.In addition to base salary, this role may be eligible for bonus, equity, and/or commission programs.Nscale may offer a competitive benefits package including medical, dental, vision, flexible paid time off, parental leave, and retirement plan participation.Salary Range$240,000 - $320,000 USDFor information on how Nscale handles candidate personal data, please see our Employee & Candidate Privacy Notice:Here.#J-*****-Ljbffr

Requirements

Bachelor’s degree in Computer Science, Engineering, or a related technical field, or equivalent practical experience. 10+ years of experience leading network, compute, infrastructure, or platform engineering teams. Experience operating within a large cloud provider, internet service provider, hyperscale data center, or similarly complex infrastructure environment. Experience leading network or compute operations teams responsible for business-critical production environments. Demonstrated experience building and scaling engineering teams and delivering complex, multi-quarter infrastructure initiatives. Strong understanding of large-scale data center, cloud, network, or compute architecture. Ability and willingness to travel up to 50%. Preferred Qualifications Experience designing or operating high-performance GPU networks using InfiniBand or Ethernet-based RoCE. Strong knowledge of networking protocols and technologies including BGP, OSPF, IS-IS, MPLS, TCP/IP, IPv4, IPv6, DNS, DHCP, VPN, and SSL. Experience with overlay networking technologies, including VXLAN and EVPN. Experience with server and GPU hardware architecture, lifecycle management, and systems management. Familiarity with system-level architecture, distributed systems, data synchronization, state management, fault tolerance, and high-availability design. Experience with infrastructure automation, scripting, network validation, and data center design. Experience implementing network monitoring, observability, and telemetry platforms. Broad experience across enterprise storage, networking, compute, and cloud infrastructure. Proven ability to resolve complex technical and organizational challenges using sound judgment and creative problem-solving. Experience influencing product roadmaps, technical priorities, and platform investment decisions. Excellent organizational, written, and verbal communication skills. Ability to influence senior stakeholders and create alignment across engineering, product, operations, vendors, and customers.

Benefits & conditions

Actual compensation may vary based on job-related factors such as skill set, experience, education, and location. In addition to base salary, this role may be eligible for bonus, equity, and/or commission programs. Nscale may offer a competitive benefits package including medical, dental, vision, flexible paid time off, parental leave, and retirement plan participation. Salary Range $240,000 - $320,000 USD For information on how Nscale handles candidate personal data, please see our Employee & Candidate Privacy Notice:Here. #J-*****-Ljbffr

About the company

Girona, España

About Nscale Nscale is the GPU cloud engineered for AI. We provide high-performance, scalable infrastructure to AI startups and large enterprise customers, reducing the complexity of developing, deploying, and operating AI workloads. At Nscale, we operate with urgency, ownership, and accountability. We value open communication, practical problem-solving, and engineering excellence as we build the infrastructure powering the future of AI.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.buscojobs.com.es

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:20 min

Identifying multi-disciplinary talent for developer experience engineering roles

Hazal Mestci +1 · Coffee With Developers

5:02 min

Mapping distributed compute paradigms to modern vehicles

Joachim Werner · LIVE

2:22 min

Introducing Skupper for application connectivity

Alex Soto Alex Soto · World Congress 2024

1:34 min

The pros and cons of campus-wide IP authentication

Christoph Eicke Christoph Eicke · World Congress 2025

3:50 min

Queues in TCP stacks and continuous network connections

Clemens Vasters Clemens Vasters · World Congress 2022

3:05 min

Exploring microcontrollers and communication protocols for amateur hardware

Philipp-Alexander Blum · LIVE

Videos

See all

Related articles

See all