Linux Infrastructure Engineer (Bare Metal, Storage & AI Factory Infrastructure)

Unknown

Remote regions

Global

Benefits

Similar Jobs

See all

Infrastructure Engineering:

  • Manage large-scale Linux deployments from hardware layer up, including BIOS/UEFI, RAID, firmware, and diagnostics.
  • Operate Bare Metal as a Service (BMaaS) platforms and lifecycle management for servers in data centers.
  • Troubleshoot complex issues across OS, hardware, storage, and networking for mission-critical production environments.

AI Factory & GPU:

  • Deploy and manage GPU-accelerated infrastructure with NVIDIA A100, H100, H200, and DGX systems.
  • Support GPU clusters, AI training environments, and HPC workloads with scheduling and networking optimization.
  • Implement high-bandwidth, low-latency designs and ecosystem technologies like CUDA, NCCL, and GPUDirect Storage.

Enterprise Storage & Networking:

  • Administer advanced Linux storage including LVM, XFS, NFS, iSCSI, Fibre Channel, and Ceph clusters for capacity and performance.
  • Configure and troubleshoot high-performance networking such as 100G/400G Ethernet, RoCE, RDMA, and spine-leaf architectures.
  • Ensure redundant, high-availability systems with monitoring and disaster recovery practices.

Unknown

The company is seeking a senior Linux infrastructure engineer with expertise in bare metal, storage, and AI Factory platforms. The size, employees, and culture are not specified.

Apply for This Position