About the role
CoreWeave is seeking an experienced, people-focused Engineering Manager to lead our MetalDev RAS (Reliability, Availability & Serviceability) team within Hardware Engineering Dev (Metal Dev). In this first-line management role, you will build, coach, and grow a team of infrastructure and site reliability engineers responsible for the reliability, availability, and serviceability of the services that manage CoreWeave's bare-metal infrastructure at scale. You will own both the health of your team,
Requirements
3+ years of engineering management experience leading and growing teams of software / infrastructure / SRE engineers (first-line management), on top of a strong individual-contributor background. 7+ years of combined experience in cloud operations, site reliability engineering (SRE), infrastructure, or related technical roles. Strong understanding of cloud platforms and K8S, and of cloud and bare-metal infrastructure. Solid grounding in incident management practices and frameworks (e.g., ITIL, S
About the company
CoreWeave
CoreWeave is a cloud provider specializing in an AI-native platform built to power complex AI workloads. They offer GPU and CPU compute, storage, and infrastructure control solutions, positioning themselves as the essential cloud for AI innovation and significantly reducing total cost of ownership.
View company page →More from CoreWeave
Senior Systems Engineer, Virtualization
LinuxKubernetesGoRustC
$182k–242k
Onsite · Livingston, NJ / New York, NY / Sunnyvale, CA / Bellevue, WA
Senior Systems Engineer, Virtualization
$182k–242k
LinuxKubernetesGoRustC
Staff Storage Engineer
RustC++GoKubernetesS3ClickHousePrometheusGrafana
$207k–303k
Onsite · Sunnyvale, California
Staff Storage Engineer
$207k–303k
RustC++GoKubernetesS3ClickHousePrometheusGrafana
Staff Software Engineer, Data Infrastructure Services
NATSApache KafkaKubernetesLinuxGoPython
$207k–275k
Onsite · Sunnyvale, CA / Bellevue, WA
Staff Software Engineer, Data Infrastructure Services
$207k–275k
NATSApache KafkaKubernetesLinuxGoPython
Sr. Infrastructure Engineer
GoPythonPrometheusGrafanaKubernetesAWSGCP
$153k–242k
Onsite · Livingston, NJ / New York, NY / Sunnyvale, CA / Bellevue, WA
Sr. Infrastructure Engineer
$153k–242k
GoPythonPrometheusGrafanaKubernetesAWSGCP
Senior Software Engineer- Billing Product
Go
$182k–242k
Onsite · New York, NY / Bellevue, WA
Senior Software Engineer- Billing Product
$182k–242k
Go
Similar roles
Senior Manager of Site Reliability Engineering
$187k–243k
Senior Manager of Site Reliability Engineering
$187k–243k
Engineering Manager, Cloud Infrastructure
$240k–300k
Engineering Manager, Cloud Infrastructure
$240k–300k
Site Reliability Engineer (SRE)
€55k–68k
Site Reliability Engineer (SRE)
€55k–68k
Senior Site Reliability Engineer, Fleet Management
$127k–249k
Senior Site Reliability Engineer, Fleet Management
$127k–249k
Software Engineer, Infrastructure
$164k–227k
Software Engineer, Infrastructure
$164k–227k