Senior AI Compute Cluster Operations Engineer
Save this job and keep your search organized
Create a free account to save jobs, create alerts and return to this listing from your dashboard.
Cerebras Systems seeks an AI Cluster Operations Engineer to manage the world's largest AI compute clusters, including the Wafer-Scale Engine. You will ensure health, performance, and availability of infrastructure, maximize capacity, and support AI initiatives.
The role requires Linux expertise, Docker/Kubernetes, and experience with monitoring, automation, and distributed systems. You will work with a fast-paced, on-call team to deliver reliable, scalable platforms for cutting-edge AI workloads.