ForgeApply · Job listing
Senior Systems Engineer, Virtualization
Coreweave
Apply in about a minute — without sacrificing quality.
ForgeApply autofills this application and tailors your resume to this exact posting. You review everything before it's sent. Free trial, no card required.
About this role
CoreWeave is The Essential Cloud for AI™. Built for pioneers by pioneers, CoreWeave delivers a platform of technology, tools, and teams that enables innovators to build and scale AI with confidence. Trusted by leading AI labs, startups, and global enterprises, CoreWeave combines superior infrastructure performance with deep technical expertise to accelerate breakthroughs and turn compute into capability. Founded in 2017, CoreWeave became a publicly traded company (Nasdaq: CRWV) in March 2025. Learn more at www.coreweave.com .
What You'll Do:
HAVOCK builds the software stack that bridges AI workloads and bare metal. We own the operating system, virtualization, runtime, and hardware interfaces that allow thousands of GPU servers to securely execute customer workloads at hyperscale.
This team focuses on the execution layer beneath Kubernetes. We build the systems that provide strong workload isolation, efficient GPU sharing, and high-performance execution across containers and lightweight virtual machines. Our work spans Linux, KVM, container runtimes, GPU drivers, and Kubernetes, ensuring customers can safely run demanding AI workloads on shared infrastructure without sacrificing performance.
About the role:
As a Senior Software Engineer on HAVOCK's Runtime & Virtualization team, you'll design and build the execution environment that powers CoreWeave's AI platform. This is fundamentally a Linux systems engineering role where you'll work across the Linux kernel, KVM/QEMU, container runtimes, GPU drivers, and Kubernetes to solve problems that don't have off-the-shelf solutions.
You'll develop secure sandboxed runtimes for GPU workloads, extend virtualization technologies to support new hardware capabilities, optimize the interaction between Linux, hypervisors, and NVIDIA GPUs, and build the tooling that helps engineers understand what's happening across the entire software stack.
The work spans multiple abstraction layers. One day you might be debugging a kernel memory-management issue affecting VFIO device passthrough; the next you might be improving container startup latency, extending KubeVirt to support new GPU workflows, or building eBPF tooling to diagnose production networking and scheduling problems. We value engineers who enjoy understanding how systems behave from the hardware up rather than treating infrastructure as a black box.
Some of what you'll work on:
• Design secure execution environments using containerd, runc, gVisor, Kata Containers, KubeVirt, and KVM/QEMU.
• Build GPU-aware runtime infrastructure supporting VFIO, Kata, NVIDIA GPU Operator, and PCIe passthrough for multi-tenant AI workloads.
• Improve Linux kernel and hypervisor performance through optimization of scheduling, memory management, I/O, NUMA locality, and virtualization primitives.
• Debug complex interactions across Linux, KVM, GPU drivers, firmware, and Kubernetes when workloads don't behave as expected.
• Develop observability and debugging tooling using eBPF, perf, tracepoints, and kernel tracing infrastructure.
• Improve container and VM startup performance, resource isolation, and runtime efficiency for latency-sensitive AI inference and training workloads.
• Extend virtualization infrastructure supporting virtio devices, IOMMU, SR-IOV, mediated devices, nested virtualization, and hardware passthrough.
• Profile production systems and build performance analysis tooling to identify bottlenecks across kernels, hypervisors, container runtimes, storage, networking, and GPUs.
• Collaborate with security, platform, networking, and GPU infrastructure teams to define the next generation of runtime isolation and workload execution.
Who You Are:
• 5+ years building production systems software, platform infrastructure, virtualization, or Linux-based distributed systems.
• Strong Linux systems knowledge, including namespaces, cgroups, scheduling, memory management, filesystems, networking, and process lifecycle.
• Experience with virtualization technologies such as KVM, QEMU, VFIO, virtio, Kata Containers, KubeVirt, Firecracker, or gVisor.
• Experience building or operating Kubernetes platforms and container runtimes at scale.
• Strong systems programming skills in Go, Rust, C/C++, or a combination thereof.
• Comfortable debugging production failures that span hardware, operating systems, container runtimes, virtualization, and distributed infrastructure.
• Experience profiling and optimizing system performance using tools such as perf, eBPF, ftrace, bpftrace, flame graphs, or crash analysis.
Preferred:
• Linux kernel development or kernel module experience.
• Experience debugging kernel panics, crash dumps, memory corruption, or driver issues using kdump, crash, drgn, or gdb.
• Familiarity with NVIDIA GPU drivers, GPU Operator, CDI, and GPU virtualization.
• Experience contributing to Linux, KVM, QEMU, Kata Containers, gVisor, containerd, or Kubernetes.
• Understanding of PCIe, IOMMU, DMA, NUMA, interrupts, and modern server hardware architecture.
• Experience building systems that safely execute untrusted workloads in shared environments.
The base salary range for this role is $182,000 to $242,000. The starting salary will be determined based on job-related knowledge, skills, experience, and market location. We strive for both market alignment and internal equity when determining compensation. In addition to base salary, our total rewards package includes a discretionary bonus, equity awards, and a comprehensive benefits program (all based on eligibility). What We Offer
The range we’ve posted represents the typical compensation range for this role. To determine actual compensation, we review the market rate for each candidate which can include a variety of factors. These include qualifications, experience, interview performance, and location.
In addition to a competitive salary, we offer a variety of benefits to support your needs. The benefits below reflect our US-based offer
Salary insight
The midpoint of this range ($212k) is about 20% above the median disclosed salary for New York roles listed on ForgeApply ($176k across 4,656 jobs).
Based on live postings with disclosed pay on ForgeApply; refreshed daily. Not an estimate of this employer's offer.
Ready to apply to Coreweave?
Apply in about a minuteSimilar jobs
- Infrastructure Engineer - Virtualization — Tensorwave · Remote
- Infrastructure Engineer - Virtualization — Tensorwave · Las Vegas, Nevada
- Senior Software Engineer II, Virtualization — Digitalocean98 · Seattle
- Senior Staff Infrastructure Engineer - Virtualization — Tensorwave · Remote
- Systems Engineer, VDI Platform — Coreweave · Livingston, NJ / New York, NY / Sunnyvale, CA / San Francisco, CA / Bellevue, WA
- Senior Systems Engineer — Barbaricum · Falls Church, VA
- Senior Systems Engineer — Crusoe · San Francisco, CA - US
- Senior Systems Engineer — 2k · Novato, California, United States
More like this: More jobs at Coreweave · Browse all jobs