Linux Infrastructure Engineer
Uvation
Job Overview
We are seeking a highly experienced Senior Linux Infrastructure Engineer with deep expertise in Linux administration, bare metal infrastructure, enterprise storage, and next-generation AI Factory / GPU infrastructure platforms . This role is focused on designing, deploying, operating, and troubleshooting large-scale Linux-based infrastructure that powers both traditional enterprise workloads and modern AI/ML environments.
This is not a DevOps-focused role . We already have a dedicated DevOps team and are looking for an engineer with extensive hands-on experience in Bare Metal as a Service (BMaaS), GPU infrastructure, high-performance storage, data center operations, and enterprise Linux platforms .
The ideal candidate will have experience building and managing infrastructure from the hardware layer up, including servers, networking, storage, GPU clusters, and AI-ready platforms. They should be comfortable working with high-performance computing (HPC), AI Factory environments, and large-scale Linux deployments where performance, reliability, and operational excellence are critical.
Key Responsibilities & Required Skills
Linux & Bare Metal Infrastructure
- Expert-level Linux administration (Ubuntu required; Red Hat and SUSE preferred)
- Deep expertise in bare metal server deployment, architecture, provisioning, and lifecycle management
- Experience operating Bare Metal as a Service (BMaaS) platforms and large-scale infrastructure environments
- Strong understanding of server hardware, including:
- BIOS/UEFI
- RAID controllers
- Firmware management
- iLO/iDRAC/IPMI
- NICs and SmartNICs
- HBA cards
- Hardware diagnostics and troubleshooting
- Experience designing, implementing, and supporting enterprise Linux infrastructure at scale
AI Factory & GPU Infrastructure
- Experience deploying and managing GPU-accelerated infrastructure for AI/ML workloads
- Understanding of NVIDIA GPU technologies including:
- A100, H100, H200, B200, or equivalent GPU platforms
- NVIDIA DGX and OEM GPU servers
- GPU provisioning and lifecycle management
- GPU monitoring and performance optimization
- Knowledge of AI Factory architecture and infrastructure requirements
- Experience supporting GPU clusters, AI training environments, and high-performance computing (HPC) workloads
- Understanding of:
- GPU resource allocation and scheduling
- Multi-GPU systems
- GPU networking requirements
- High-bandwidth, low-latency infrastructure design
- Familiarity with NVIDIA ecosystem technologies such as:
- CUDA
- NCCL
- GPUDirect Storage
- NVIDIA Fabric Manager
- NVIDIA Base Command (preferred)
Enterprise Storage & Data Platforms
- Advanced Linux storage administration:
- LVM
- XFS, EXT4
- NFS
- iSCSI
- Fibre Channel SAN
- Multipath I/O
- Strong hands-on experience with Ceph , including:
- Cluster architecture
- MON, OSD, MDS
- RBD, CephFS, RGW
- Capacity planning
- Performance tuning
- Failure recovery
- Experience with high-performance AI storage platforms such as:
- WEKA
- VAST Data
- Dell PowerScale
- Pure Storage FlashBlade
- NetApp
- Understanding of:
- NVMe-over-Fabrics (NVMe-oF)
- RDMA
- GPUDirect Storage
- Parallel file systems
- AI data pipelines
Networking & Infrastructure
- Strong networking knowledge:
- Bonding
- VLANs
- Routing
- MTU optimization
- DNS
- DHCP
- Experience with high-performance data center networking:
- 100G/200G/400G Ethernet
- RoCE
- RDMA
- Spine-Leaf architectures
- Familiarity with NVIDIA Spectrum-X, Mellanox/NVIDIA ConnectX adapters, or equivalent technologies
- Strong understanding of Layer 2 and Layer 3 infrastructure design and troubleshooting
Operations & Reliability
- Experience with high availability, clustering, and disaster recovery
- Strong troubleshooting skills across:
- Linux operating systems
- Hardware platforms
- GPU infrastructure
- Networking
- Enterprise storage
- Experience supporting mission-critical production environments
- Bash and Python scripting for automation and operational efficiency
- Experience creating operational documentation, runbooks, and infrastructure standards
- Understanding of AI infrastructure design and reference architectures
- AI cloud integration for workloads
- SOP and runbook development and maintenance
- Incident, problem, and capacity management
- Business continuity and disaster recovery planning for AI workloads
- Proactive risk identification and mitigation to avoid business impact
Nice to Have
- Kubernetes infrastructure (especially AI/ML and GPU integration)
- KVM, VMware, OpenShift Virtualization, or similar virtualization platforms
- Ansible automation
- NVIDIA Base Command Manager
- Slurm or HPC workload schedulers
- Observability and monitoring platforms (Prometheus, Grafana, OpenTelemetry)
- Data Center Infrastructure Management (DCIM) tools
- IPAM solutions
- AWS, Azure, or hybrid cloud exposure
We Are Not Looking For
- Candidates whose experience is primarily CI/CD pipeline engineering
- Engineers focused mainly on Terraform, GitOps, or application delivery pipelines
- Cloud-only administrators with limited bare metal, storage, or hardware experience
- Professionals whose primary expertise is software development rather than infrastructure engineering
Ideal Candidate
Someone who has spent years designing, building, and operating enterprise Linux environments, large-scale bare metal infrastructure, storage platforms, and modern AI Factory environments. The ideal candidate understands how to deploy and manage GPU-enabled infrastructure, BMaaS platforms, enterprise storage, and high-performance networking while solving complex operating system, hardware, storage, and AI infrastructure challenges. DevOps experience is a plus, but deep Linux, infrastructure, storage, BMaaS, and AI Factory expertise is the primary requirement.
$1000 per year
...(USA & Japan, 2024–2025) and a Top-5 Company for Work-From-Anywhere Jobs (FlexJobs, 2025). We are looking for an Infrastructure Security Engineer . Your main tasks will be: Design, implement, and maintain Cloudflare Zero Trust architecture (Access,...- ...programmable blockchain platform. We’re hiring a Staff Engineer for the Developer Infrastructure team within Platform Engineering at Ava Labs. Platform... ...pragmatically relieve bottlenecks. – Working fluency with Linux, containers, and infrastructure-as-code. Preferred...
$1000 per year
...CV, keep reading. About the role: At Camunda, our CI/CD infrastructure platform is treated as an internal product. Our mission is to build a platform that offers a Golden Path, enabling 150+ engineers across 20+ teams to deliver with world-class velocity and uncompromising...$2000 per year
...This is a general track for applications to any team at Canonical that works with the Linux kernel, across all seniority levels. Apply here if you are an exceptional software engineer who wants to work on both stable and cutting edge Linux kernels for Ubuntu and its wider...$2000 per year
...Work across the full Linux stack from kernel through GUI to optimise Ubuntu, the world’s most widely used Linux desktop and server,... ...across PC and IoT technologies. Our teams partner with specialist engineers from major silicon companies to integrate next-generation...- ...s cloud environments, networks, compute platforms, and core infrastructure across AWS and GCP. Strengthen configuration, access controls... ..., recovery, business continuity, and secure operations with engineering teams. WHAT WE LOOK FOR – Degree in Computer Science,...
$2000 per year
...Ubuntu Linux, already the most popular Linux distribution in the world, is looking to increase its adoption even further by expanding... ...candidate will be able to prove a strong aptitude for software engineering at the hardware level. While direct experience with the Linux...$2000 per year
...enterprise initiatives such as public cloud, data science, AI, engineering innovation, and IoT. Our customers include the world's leading... ...junior professionals into the Canonical Kernel Team, to work on the Linux kernel for Ubuntu. If you’ve enjoyed operating systems in your...- ...safer and more secure. The Data Platform team builds and owns highly available, scalable data infrastructure for TRM’s products and services. As a Senior Software Engineer on Data Infrastructure (RDBMS), you will develop, operate, and scale the relational database...
$2000 per year
...such as public cloud, data science, AI, engineering innovation, and IoT. Our customers... ...also central to the health of critical infrastructure across the globe. As Ubuntu has been embraced... ...and spoken English Experience with Linux (Debian or Ubuntu preferred)...- ...with enterprise architecture, security, engineering, and deployment standards.... ...in enterprise GIS implementation, cloud infrastructure, and Web-GIS application development.... ...Kubernetes, Terraform, Windows Server, and Linux. ~ Strong software development experience...
- ...Wanted: Colonist DevOps Engineer Working Hours: We work asynchronously and value outcomes... ...a single game. You will own the infrastructure behind every game session: clusters, deploys... ...your work. – You are comfortable with Linux, Docker, and a CI system you did not...
$2000 per year
...This is the general track for Engineering Director at Canonical, apply here if you are confident... ...and Golang C / C++ / Rust Data infrastructure HTML / CSS / JavaScript / Typescript... ...set high expectations Outstanding Linux based software engineering track record...$2000 per year
...such as public cloud, data science, AI, engineering innovation, and IoT. Our customers... ...Integrate new tools into our security infrastructure, pipelines, and processes Achieve and... ...security certifications Extend and enhance Linux cryptographic components to meet...$2000 per year
...such as public cloud, data science, AI, engineering innovation and IoT. Our customers... ...Container, virtualisation and cloud infrastructure have become essentials of modern software... ...great potential as a new hypervisor for Linux. We are building a team to work on this...- ...ABOUT THE TEAM The Platform Engineering team builds the tools and services that accelerate... ..., implement and maintain the build infrastructure for a large Typescript and NodeJS codebase... ...areas – A thorough understanding of Linux, Docker and CI tools like Github Actions...
$2000 per year
...such as public cloud, data science, AI, engineering innovation and IoT. Our customers... ...the globe Work on shared tools and infrastructure for performance measurement, analysis and... ...presentation skills Experience with Linux (Debian or Ubuntu preferred) Excellent...$2000 per year
...enterprise initiatives such as public cloud, data science, AI, engineering innovation and IoT. Our customers include the world's leading public... .... These roles require extensive personal experience with Linux - the more different versions of Linux the better! Location...$2000 per year
...initiatives such as public cloud, data science, AI, engineering innovation, and IoT. Our customers include the world'... ...Engineering for ... ...a fast-paced engineering role in Linux-based software-defined infrastructure and applications, covering all layers of the stack,...$2000 per year
...such as public cloud, data science, AI, engineering innovation, and IoT. Our customers... ...revolutionise open source application and infrastructure operations. We want to transform the... ...computing, and an interest in the entire Linux stack - from kernel to networking to...$2000 per year
...enterprise initiatives such as public cloud, data science, AI, engineering innovation, and IoT. Our customers include the world's leading... ...is well-known as a developer favourite and enterprise-friendly Linux. Our web services and system utilities are often written in Python...$2000 per year
...such as public cloud, data science, AI, engineering innovation and IoT. Our customers... ...solutions for public cloud and private infrastructure. As a software engineer on the team,... ...Develop your understanding of the entire Linux stack, from kernel, networking, and storage...- ...others. We are looking for a Backend Engineer to join our Core team and help build... ...of production systems. Contribute to infrastructure improvements, including CI/CD, observability... .... ~ Experience working with Linux and Docker. ~ Ability to write and maintain...
- ...We are seeking a skilled and experienced Azure DevOps Engineer with a strong background in Linux administration to join our dynamic team. The ideal... ...responsible for managing and optimizing our Azure cloud infrastructure, ensuring seamless CI/CD pipelines, and maintaining...
$2000 per year
...such as public cloud, data science, AI, engineering innovation, and IoT. Our customers... ...systems, and an interest in the entire Linux stack - from kernel to networking to virtualization... ...think rigorously about application and infrastructure reliability Shape high quality open...- ...As an Engineering Manager on Coder’s Core Workspaces team, you’ll lead engineers building... ...experience with AWS, Kubernetes, Docker, or Linux. – Open-source contributions or... ...– Frontend: TypeScript, React – Infrastructure: AWS, Kubernetes – Observability: Prometheus...
- ...MongoDB clusters. You are experienced with modern infrastructure deployment automations or with traditional Linux systems administration, operations, and package... ...'s pioneers in open source with intelligent engineers at every level from engineer to CTO and CEO level...
$2000 per year
...ecosystem. This is your chance to be a part of that as a Community Engineer at Canonical . We are building community management at scale.... ...person who is passionate about open source software, Linux, and sustainable community building. In this role, you will...$2000 per year
...such as public cloud, data science, AI, engineering innovation, and IoT. Our customers... ...Creating automated testing approaches and infrastructure for validating reliability, performance... ...fundamentals ~ Solid understanding of the Linux system architecture ~ Complex...$2000 per year
...enterprise initiatives such as public cloud, data science, AI, engineering innovation, and IoT. Our customers include the world's leading... ...that there is an opportunity to rethink the foundations of future Linux systems with Rust as a central driver of change in everything...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Linux Infrastructure Engineer. Be the first to apply!
