Mirantis is building a specialized European AI Infrastructure & Platform Operations team dedicated to managing large-scale environments powered by NVIDIA GPUs, high-performance networking, and Kubernetes. You will play a critical role in ensuring the availability, performance, and operational stability of the infrastructure that drives modern AI workloads.
Key responsibilities
- Monitor, operate, and support production AI infrastructure platforms.
- Investigate and resolve infrastructure, networking, hardware, and platform-related incidents.
- Support NVIDIA GPU infrastructure and associated platform services.
- Monitor and troubleshoot Kubernetes-based environments.
- Collaborate with engineering teams and hardware vendors to resolve technical issues.
Requirements
- 3+ years of experience in infrastructure operations, platform operations, or site reliability engineering.
- Strong Linux administration and troubleshooting skills.
- Good understanding of networking concepts and experience diagnosing infrastructure issues.
- Working knowledge of Kubernetes in production environments.
- Excellent analytical, problem-solving, and communication skills.
What we offer
- Work with some of the most advanced AI infrastructure environments in production today.
- Gain exposure to NVIDIA GPU technologies, Kubernetes platforms, and high-performance networking.
- Help define how next-generation AI infrastructure is operated and supported.
- Join a growing organization investing heavily in AI infrastructure and platform services.