Similar Jobs
See allAI Infrastructure & Platform Operations Engineer
Mirantis
Europe
Linux
Kubernetes
Networking
Senior AI Infrastructure & Platform Operations Engineer
Mirantis
Europe
Linux
Kubernetes
Networking
Senior Software Systems Engineer (Storage)
Mirantis
US
Kubernetes
Linux
Infrastructure As Code
HPC Network Engineer
Mirantis
Global
Networking
Linux
InfiniBand
Senior DevOps Engineer (Storage)
Mirantis
US
Kubernetes
Linux
Terraform
Infrastructure Engineering:
- Manage large-scale Linux deployments from hardware layer up, including BIOS/UEFI, RAID, firmware, and diagnostics.
- Operate Bare Metal as a Service (BMaaS) platforms and lifecycle management for servers in data centers.
- Troubleshoot complex issues across OS, hardware, storage, and networking for mission-critical production environments.
AI Factory & GPU:
- Deploy and manage GPU-accelerated infrastructure with NVIDIA A100, H100, H200, and DGX systems.
- Support GPU clusters, AI training environments, and HPC workloads with scheduling and networking optimization.
- Implement high-bandwidth, low-latency designs and ecosystem technologies like CUDA, NCCL, and GPUDirect Storage.
Enterprise Storage & Networking:
- Administer advanced Linux storage including LVM, XFS, NFS, iSCSI, Fibre Channel, and Ceph clusters for capacity and performance.
- Configure and troubleshoot high-performance networking such as 100G/400G Ethernet, RoCE, RDMA, and spine-leaf architectures.
- Ensure redundant, high-availability systems with monitoring and disaster recovery practices.
Unknown
The company is seeking a senior Linux infrastructure engineer with expertise in bare metal, storage, and AI Factory platforms. The size, employees, and culture are not specified.