← All roles

Software Engineer, Inference AI/ML

CoreWeave ·Sunnyvale, CA / Bellevue, WA ·Onsite 31d ago
PythonGoC++vLLMKubernetesLinuxGitPyTorchTensorFlowCUDAGrafanaPrometheusOTEL

About the role

Join the Inference team to ship production features that improve latency, reliability, and cost for model serving on our GPU platform. As an IC1, you will implement well-scoped changes, learn operational practices, and grow quickly with mentorship. Responsibilities include implementing features in Python/Go/C++ for model-serving services like Triton, vLLM, TensorRT-LLM, and Ray Serve; writing tests and design docs; adding metrics and dashboards; following on-call runbooks; and contributing to b,

Requirements

BS/MS in CS, EE, or related field, or equivalent practical experience. Foundations in data structures, algorithms, and networked services. Experience with Python or Go (C++ a plus) and Linux fundamentals; Git/CI basics. Exposure to containers and Kubernetes. Curiosity about GPU inference concepts. Preferred: internship or project deploying a microservice or ML inference demo; coursework with PyTorch or TensorFlow; familiarity with Grafana/Prometheus/OpenTelemetry.

About the company

CoreWeave

CoreWeave is a cloud provider specializing in an AI-native platform built to power complex AI workloads. They offer GPU and CPU compute, storage, and infrastructure control solutions, positioning themselves as the essential cloud for AI innovation and significantly reducing total cost of ownership.

View company page →
$92k–135k
Onsite Sunnyvale, CA / Bellevue, WASunnyvale, USBellevue, US
Posted 31d ago
Apply for this role
Opens coreweave.com ↗
About the company
CoreWeave
https://coreweave.com

CoreWeave is a cloud provider specializing in an AI-native platform built to power complex AI workloads. They offer GPU and CPU compute, storage, and infrastructure control solutions, positioning themselves as the essential cloud for AI innovation and significantly reducing total cost of ownership.

View company page →