Staff Cloud SRE for AI/ML Platform & GPU Compute

Posted 2 hours 56 minutes ago by OpenDigital Limited

Permanent
Full Time
Other
London, United Kingdom
Job Description

Wayve is seeking a founding Staff Cloud Site Reliability Engineer to shape the reliability of large-scale AI systems and GPU compute infrastructure. You will build and scale the reliability foundations of our AI cloud platform, including model development and GPU compute environments.

This is a founding SRE role situated at the intersection of AI research, cloud infrastructure, and operations. You will define frameworks, automate standards, and ensure scalable, secure deployments with a bias