Senior AI Compute Cluster Operations Engineer

5 days ago

, Canada Cerebras Full-time

Cerebras Systems seeks an AI Cluster Operations Engineer to manage the world's largest AI compute clusters, including the Wafer-Scale Engine. You will ensure health, performance, and availability of infrastructure, maximize capacity, and support AI initiatives.

The role requires Linux expertise, Docker/Kubernetes, and experience with monitoring, automation, and distributed systems. You will work with a fast-paced, on-call team to deliver reliable, scalable platforms for cutting-edge AI workloads.