Staff DevOps Engineer with AI Focus
7 days ago
Winnipeg, Manitoba, Canada
Nexxa.ai
Full-time
Free with email or Google
Save this job and keep your search organized
Create a free account to save jobs, create alerts and return to this listing from your dashboard.
Free with email or Google
Innovate infrastructure solutions as a Staff DevOps Engineer, specializing in AI at Nexxa. Shape systems for emerging industrial applications and enhance operational reliability.
Nexxa is seeking a Staff DevOps Engineer who will take ownership of infrastructure supporting AI applications in heavy industries. You will ensure that production ML workloads perform optimally in real-world settings. Collaborate closely with product engineering teams to design CI/CD systems and maintain on-prem and cloud deployments for mission-critical operations.
Key Responsibilities: • Design and manage CI/CD pipelines for AI and product teams • Build and maintain infrastructure-as-code for environments • Architect Kubernetes platforms with GPU workload scheduling • Define metrics and logging practices across systems • Enforce reliability protocols and incident management
Requirements: • 6+ years in relevant DevOps or infrastructure roles • Extensive experience with AWS, GCP, or Azure • Proficient in Kubernetes and infrastructure-as-code tools • Strong background in observability stack management • Excellent automation skills in Python, Go, or Bash
Drive AI infrastructure innovations that empower Nexxa’s industrial solutions.
Nexxa is seeking a Staff DevOps Engineer who will take ownership of infrastructure supporting AI applications in heavy industries. You will ensure that production ML workloads perform optimally in real-world settings. Collaborate closely with product engineering teams to design CI/CD systems and maintain on-prem and cloud deployments for mission-critical operations.
Key Responsibilities: • Design and manage CI/CD pipelines for AI and product teams • Build and maintain infrastructure-as-code for environments • Architect Kubernetes platforms with GPU workload scheduling • Define metrics and logging practices across systems • Enforce reliability protocols and incident management
Requirements: • 6+ years in relevant DevOps or infrastructure roles • Extensive experience with AWS, GCP, or Azure • Proficient in Kubernetes and infrastructure-as-code tools • Strong background in observability stack management • Excellent automation skills in Python, Go, or Bash
Drive AI infrastructure innovations that empower Nexxa’s industrial solutions.