EU remote
Staff Software Engineer, Cluster Orch (SUNK)
About this role
CoreWeave is The Essential Cloud for AI™. Built for pioneers by pioneers, CoreWeave delivers a platform of technology, tools, and teams that enables innovators to build and scale AI with confidence. Trusted by leading AI labs, startups, and global enterprises, CoreWeave combines superior infrastructure performance with deep technical expertise to accelerate breakthroughs and turn compute into capability. Founded in 2017, CoreWeave became a publicly traded company (Nasdaq: CRWV) in March 2025.
Learn more at www.coreweave.com . We're proud to be a Living Wage accredited Employer. What You'll Do: CoreWeave’s Cluster Orchestration team builds and operates the Kubernetes-native foundation that powers AI training and inference at scale. We eliminate infrastructure bottlenecks and create next-generation orchestration capabilities—including SUNK (Slurm on Kubernetes) and beyond—ensuring demanding workloads run seamlessly, reliably, and efficiently across massive GPU clusters.
About the role: As a Staff Software Engineer, Cluster Orch (SUNK), you will act as a principal technical leader shaping the long-term architecture and strategy for CoreWeave’s orchestration platform. You will define the technical direction, own critical components of our orchestration layers and managed services, and drive major cross-organisational initiatives in scheduling, multi-tenant quota enforcement, and hyperscale scaling.
In this high-impact role, you will establish organisation-wide best practices for platform reliability and observability, resolve complex distributed bottlenecks under intense demand, and mentor senior engineers across teams to elevate technical standards. Who You Are: Bachelor’s degree in Computer Science, Engineering, or a related technical field (or equivalent practical experience). 8+ years of professional software engineering experience with a proven track record of designing, operating, and scaling large-scale distributed systems in production environments.
Advanced software development proficiency in Go and strong distributed systems design principles. Deep technical expertise in Kubernetes internals, Slurm schedulers, or cloud-native development. Demonstrated experience setting technical direction and successfully influencing cross-team architecture and platform goals. Proven ability to mentor senior engineers, review complex technical designs, and elevate organisational operational standards.