Jobiglo

No results.

Linux Infrastructure Engineer – Bare Metal, Storage & AI Factory

Uvation

Remote
Contrat Remote Senior 🇬🇧 English
Ubuntu Red Hat SUSE BMaaS BIOS UEFI RAID iLO iDRAC IPMI SmartNIC HBA NVIDIA DGX CUDA NCCL GPUDirect Storage NVIDIA Fabric Manager NVIDIA Base Command GPU provisioning GPU monitoring HPC AI Factory Enterprise storage Data center operations

Job description

About the role

We are looking for a senior Linux Infrastructure Engineer to design, deploy, operate and troubleshoot large‑scale Linux‑based environments that support both traditional enterprise workloads and modern AI/ML pipelines. This position focuses on bare‑metal, storage and GPU‑accelerated AI Factory infrastructure, not on DevOps tooling.

Key responsibilities

  • Architect and manage bare‑metal server fleets, including provisioning, lifecycle management and hardware diagnostics.
  • Operate and support BMaaS platforms and high‑performance storage solutions across data‑center sites.
  • Deploy, configure and maintain GPU clusters for AI/ML workloads, handling provisioning, monitoring and performance tuning.
  • Collaborate with networking and data‑center teams to ensure low‑latency, high‑bandwidth connectivity for HPC and AI workloads.
  • Provide 24/7 operational support, incident response and capacity planning for mission‑critical Linux services.

Required profile

  • Extensive hands‑on experience with enterprise Linux (Ubuntu required; Red Hat and SUSE a plus).
  • Deep knowledge of server hardware components such as BIOS/UEFI, RAID controllers, iLO/iDRAC, IPMI, SmartNICs and HBAs.
  • Proven track record building and managing GPU‑accelerated AI Factory environments.
  • Strong troubleshooting skills for both hardware and software layers in large‑scale deployments.

Required skills

  • Ubuntu, Red Hat, SUSE Linux administration
  • BMaaS platforms and bare‑metal provisioning
  • BIOS/UEFI, RAID, iLO, iDRAC, IPMI, SmartNIC, HBA
  • NVIDIA GPUs (A100, H100, H200, B200), DGX and OEM servers
  • CUDA, NCCL, GPUDirect Storage, NVIDIA Fabric Manager, NVIDIA Base Command
  • GPU provisioning, monitoring and performance optimization
  • High‑performance computing (HPC) and AI Factory architecture
  • Enterprise storage and data‑center operations

What we offer

  • Opportunity to work on cutting‑edge AI and HPC infrastructure.
  • Remote work flexibility.
  • Competitive contract‑based compensation.

Questions fréquentes

Le salaire n'est pas communiqué publiquement par le recruteur. Vous pouvez postuler et négocier directement avec Uvation.
Cliquez sur "Postuler maintenant" en haut de la page. Vous pouvez importer votre CV en 1 clic — Jobiglo extrait automatiquement vos informations et postule pour vous.
Le contrat proposé est un Contrat.

Why are you reporting this job?

Thank you for your report. We will review this job.

Explore further

Salaries, guides and searches in Malaysia.

Apply in 30 seconds

Enter your email to apply. An account will be created automatically.

Apply now →

By continuing, you accept our terms of use.

Already have an account? Login

A question about this job?

Ask it here: you will get the full job summary by e-mail, right away.

💬 Chat with us on Telegram

Published 1 month ago

Expires 2 weeks from now

23 views · 0 interested

Boost your chances

Upload your CV — we will match you with relevant openings.

Analyzing your CV...

Uvation