> Markdown version of [/jobs/ext/1894623-senior-systems-engineer-virtualization](https://www.wearedevelopers.com/jobs/ext/1894623-senior-systems-engineer-virtualization). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Systems Engineer, Virtualization - **Company:** Coreweave, Inc. - **Location:** United States - **Experience:** Expert - **Salary:** $182,000.0 - $242,000.0 - **Contract:** Permanent contract - **Skills:** Abstraction Layers, Artificial Intelligence, C++ (Programming Language), Profiling, Software Debugging, Linux, File Systems, Distributed Systems, Memory Management, Firmware, Hardware Interface Design, Hypervisor, Kernel-Based Virtual Machine, Linux Kernel, PCI Express, Performance Tuning, Quick EMUlator (QEMU), System Programming, Virtual Machines, Virtualization Technology, Loadable Kernel Module, Graphics Processing Unit (GPU), Build Management, Perf (Linux), AI Platforms, Kubernetes, Bare Metal, Hardware Infrastructure - **Published:** August 1, 2026 - **Apply:** https://www.dice.com/job-detail/dcd9649a-2d6c-4a68-a2f5-885072e41feb ## About the Role * 5+ years building production systems software, platform infrastructure, virtualization, or Linux-based distributed systems. * Strong Linux systems knowledge, including namespaces, cgroups, scheduling, memory management, filesystems, networking, and process lifecycle. * Experience with virtualization technologies such as KVM, QEMU, VFIO, virtio, Kata Containers, KubeVirt, Firecracker, or gVisor. * Experience building or operating Kubernetes platforms and container runtimes at scale. * Strong systems programming skills in Go, Rust, C/C++, or a combination thereof. * Comfortable debugging production failures that span hardware, operating systems, container runtimes, virtualization, and distributed infrastructure. * Experience profiling and optimizing system performance using tools such as perf, eBPF, ftrace, bpftrace, flame graphs, or crash analysis. Preferred: * Linux kernel development or kernel module experience. * Experience debugging kernel panics, crash dumps, memory corruption, or driver issues using kdump, crash, drgn, or gdb. * Familiarity with NVIDIA GPU drivers, GPU Operator, CDI, and GPU virtualization. * Experience contributing to Linux, KVM, QEMU, Kata Containers, gVisor, containerd, or Kubernetes. * Understanding of PCIe, IOMMU, DMA, NUMA, interrupts, and modern server hardware architecture. * Experience building systems that safely execute untrusted workloads in shared environments. ## Description HAVOCK builds the software stack that bridges AI workloads and bare metal. We own the operating system, virtualization, runtime, and hardware interfaces that allow thousands of GPU servers to securely execute customer workloads at hyperscale. This team focuses on the execution layer beneath Kubernetes. We build the systems that provide strong workload isolation, efficient GPU sharing, and high-performance execution across containers and lightweight virtual machines. Our work spans Linux, KVM, container runtimes, GPU drivers, and Kubernetes, ensuring customers can safely run demanding AI workloads on shared infrastructure without sacrificing performance., As a Senior Software Engineer on HAVOCK's Runtime & Virtualization team, you'll design and build the execution environment that powers CoreWeave's AI platform. This is fundamentally a Linux systems engineering role where you'll work across the Linux kernel, KVM/QEMU, container runtimes, GPU drivers, and Kubernetes to solve problems that don't have off-the-shelf solutions. You'll develop secure sandboxed runtimes for GPU workloads, extend virtualization technologies to support new hardware capabilities, optimize the interaction between Linux, hypervisors, and NVIDIA GPUs, and build the tooling that helps engineers understand what's happening across the entire software stack. The work spans multiple abstraction layers. One day you might be debugging a kernel memory-management issue affecting VFIO device passthrough; the next you might be improving container startup latency, extending KubeVirt to support new GPU workflows, or building eBPF tooling to diagnose production networking and scheduling problems. We value engineers who enjoy understanding how systems behave from the hardware up rather than treating infrastructure as a black box. Some of what you'll work on: * Design secure execution environments using containerd, runc, gVisor, Kata Containers, KubeVirt, and KVM/QEMU. * Build GPU-aware runtime infrastructure supporting VFIO, Kata, NVIDIA GPU Operator, and PCIe passthrough for multi-tenant AI workloads. * Improve Linux kernel and hypervisor performance through optimization of scheduling, memory management, I/O, NUMA locality, and virtualization primitives. * Debug complex interactions across Linux, KVM, GPU drivers, firmware, and Kubernetes when workloads don't behave as expected. * Develop observability and debugging tooling using eBPF, perf, tracepoints, and kernel tracing infrastructure. * Improve container and VM startup performance, resource isolation, and runtime efficiency for latency-sensitive AI inference and training workloads. * Extend virtualization infrastructure supporting virtio devices, IOMMU, SR-IOV, mediated devices, nested virtualization, and hardware passthrough. * Profile production systems and build performance analysis tooling to identify bottlenecks across kernels, hypervisors, container runtimes, storage, networking, and GPUs. * Collaborate with security, platform, networking, and GPU infrastructure teams to define the next generation of runtime isolation and workload execution. ## Related Videos - [Unleashing the Full Potential of the Arm Architecture – Write Once, Deploy Anywhere](https://www.wearedevelopers.com/videos/940-unleashing-the-full-potential-of-the-arm-architecture-write-once-deploy-anywhere) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [Profiling Symfony & PHP apps with Blackfire](https://www.wearedevelopers.com/videos/265-profiling-symfony-php-apps-with-blackfire) - [Playing Pong on a shoulder press machine](https://www.wearedevelopers.com/videos/100140-playing-pong-on-a-shoulder-press-machine) - [The best of two worlds - Bringing enterprise-grade Linux to the vehicle](https://www.wearedevelopers.com/videos/67-the-best-of-two-worlds-bringing-enterprise-grade-linux-to-the-vehicle) - [Docker exec without Docker](https://www.wearedevelopers.com/videos/1094-docker-exec-without-docker) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Why Upskilling And Reskilling is Important For Developers](https://www.wearedevelopers.com/magazine/428-why-upskilling-and-reskilling-is-important-for-developers) - [6 Emerging Technologies We’ll Learn About in 2025](https://www.wearedevelopers.com/magazine/381-6-emerging-technologies-we-ll-learn-about-in-2025)