📢 New: get today's jobs on our WhatsApp Channel
Jobiglo

No results.

Linux Infrastructure Engineer (Bare Metal, Storage & AI Factory)

Uvation

Remote
Contract Remote Senior 🇬🇧 English
Linux Ubuntu Red Hat SUSE Bare Metal as a Service BMaaS BIOS/UEFI RAID iLO iDRAC IPMI SmartNIC HBA NVIDIA DGX GPU provisioning GPU monitoring CUDA NCCL GPUDirect Storage NVIDIA Fabric Manager NVIDIA Base Command HPC AI Factory architecture

Job description

About the role

We are looking for a senior Linux Infrastructure Engineer to design, deploy, operate and troubleshoot large‑scale Linux‑based infrastructure that supports both traditional enterprise workloads and modern AI/ML environments. The role focuses on bare‑metal as a service, high‑performance storage, and GPU‑accelerated AI Factory platforms.

Key responsibilities

  • Architect and manage bare‑metal server provisioning, lifecycle, and BMaaS platforms.
  • Deploy, configure, and maintain enterprise Linux environments (Ubuntu, Red Hat, SUSE) at scale.
  • Design, install, and support GPU clusters for AI/ML workloads, including NVIDIA A100/H100/H200/B200 platforms.
  • Implement high‑performance storage solutions and ensure low‑latency networking for HPC and AI training.
  • Monitor hardware health, perform diagnostics, and resolve firmware, BIOS/UEFI, RAID, iLO/iDRAC, IPMI, SmartNIC and HBA issues.
  • Optimize GPU resource allocation, scheduling, and performance using CUDA, NCCL, GPUDirect Storage and NVIDIA Fabric Manager.

Required profile

  • Extensive hands‑on experience with Linux system administration on Ubuntu (mandatory) and preferably Red Hat or SUSE.
  • Deep knowledge of bare‑metal infrastructure, server hardware components and lifecycle management.
  • Proven track record deploying and supporting GPU‑accelerated AI/ML environments.
  • Strong understanding of high‑performance computing (HPC) and AI Factory architecture.

Required skills

  • Linux (Ubuntu, Red Hat, SUSE)
  • Bare Metal as a Service (BMaaS)
  • Server hardware: BIOS/UEFI, RAID, iLO/iDRAC, IPMI, SmartNICs, HBAs
  • NVIDIA GPUs (A100, H100, H200, B200, DGX, OEM servers)
  • GPU provisioning, monitoring and performance tuning
  • CUDA, NCCL, GPUDirect Storage, NVIDIA Fabric Manager, NVIDIA Base Command
  • High‑bandwidth, low‑latency networking design
  • HPC and AI/ML workload support

Questions fréquentes

Le salaire n'est pas communiqué publiquement par le recruteur. Vous pouvez postuler et négocier directement avec Uvation.
Cliquez sur "Postuler maintenant" en haut de la page. Vous pouvez importer votre CV en 1 clic — Jobiglo extrait automatiquement vos informations et postule pour vous.
Le contrat proposé est un Contract.

Why are you reporting this job?

Thank you for your report. We will review this job.

Apply in 30 seconds

Enter your email to apply. An account will be created automatically.

Apply now →

By continuing, you accept our terms of use.

Already have an account? Login

A question about this job?

Ask it here: you will get the full job summary by e-mail, right away.

💬 Chat with us on Telegram

Published 1 month ago

Expires 3 weeks from now

41 views · 0 interested

Boost your chances

Upload your CV — we will match you with relevant openings.

Analyzing your CV...

Uvation