Senior Site Reliability Engineer (Golang / Kubernetes)
As a Senior Site Reliability Engineer, you will define and measure reliability for a GPU-accelerated AI platform, owning service-level indicators and objectives while collaborating across teams.
As a Senior Site Reliability Engineer, you will define and measure reliability for a GPU-accelerated AI platform, owning service-level indicators and objectives while collaborating across teams.
Design, deploy, and operate the control plane for an AI-operated GPU cloud, ensuring automated remediation and efficient management of Kubernetes clusters optimized for GPU workloads.
As a Senior Site Reliability Engineer, you will design, develop, and maintain cloud-based AI infrastructure, ensuring reliability and performance while mentoring team members.
As a Senior Site Reliability Engineer, you will design, develop, and operate cloud-based AI solutions, ensuring the reliability and performance of container infrastructure while mentoring team members.
As a Senior Site Reliability Engineer, you will design, develop, and operate cloud-based AI solutions, ensuring the reliability and performance of container infrastructure while mentoring team members.