Senior Cloud Site Reliability Engineer
This role involves building and scaling the reliability foundations of Wayve's AI cloud platform, including the Model Development and GPU Compute platforms. The engineer will define SRE frameworks, automation, and operational standards for large-scale, distributed systems supporting AI training and inference. Key responsibilities include incident response, observability, capacity planning, and improving deployment safety across cloud infrastructure.