Wayve is seeking a Staff Cloud Site Reliability Engineer (AI) to build and scale the reliability foundations of our AI cloud platform, including the Model Development Platform and GPU Compute, with a focus on resilient, efficient, and scalable model development infrastructure.
This London-based role offers hybrid work (2 days in the office) and requires owning SRE practices, on-call rotations, monitoring, and automation across large GPU-backed clusters to accelerate training and deployment.
#J-18808-Ljbffr…
