Uvation is seeking a highly experienced Senior Linux Infrastructure Engineer to design, deploy, and manage large-scale bare metal, storage, and AI Factory infrastructure. This role focuses on the hardware and OS layers, supporting high-performance GPU clusters and enterprise-grade data platforms.
Key responsibilities
- Administer enterprise Linux environments with a focus on Ubuntu, Red Hat, and SUSE.
- Manage bare metal server lifecycles, including provisioning, firmware, and hardware diagnostics.
- Deploy and optimize GPU-accelerated infrastructure for AI/ML workloads, including NVIDIA platforms.
- Architect and maintain high-performance storage solutions such as Ceph, NVMe-oF, and parallel file systems.
- Support high-speed data center networking, including RoCE, RDMA, and spine-leaf architectures.
Requirements
- Expert-level knowledge of Linux administration and bare metal infrastructure operations.
- Deep understanding of server hardware, including BIOS, UEFI, RAID, and IPMI management.
- Proven experience with GPU provisioning, monitoring, and performance tuning in AI environments.
- Strong background in enterprise storage technologies and high-performance networking.
- Proficiency in Bash and Python for infrastructure automation and operational efficiency.
What we offer
- Opportunity to work on cutting-edge AI Factory and GPU-accelerated infrastructure.
- Engagement in complex, large-scale hardware and storage architecture projects.
- A fully remote work environment focused on technical excellence and infrastructure stability.
- Collaboration with a team dedicated to high-performance computing and mission-critical systems.