3 of 3 Slurm Workload Manager Jobs in the UK

Scientific Computing Engineer - Vernalis

Hiring Organisation
Vernalis
Location
Cambridgeshire, England, United Kingdom
configurations for the Enterprise IT team. Key Responsibilities HPC Management: Oversee the health and performance of our on-premises HPC cluster ( Bright Computing/SLURM ). Optimise job scheduling and resource allocation across mixed CPU and GPU nodes for computational chemistry and ML applications. DevOps & Container Orchestration: Lead … Linux (RHEL) and SELinux security policies. DevOps: Strong experience with GitLab CI/CD and container orchestration (Docker Swarm). HPC Stack: Proficiency in SLURM workload management and Bright Computing cluster tools. Storage & Identity: Experience with NFS, enterprise storage (EMC2), and managing UID/GID/SID mappings ...

Senior Solutions Engineer

Hiring Organisation
LJB & Co
Location
City of London, London, United Kingdom
Employment Type
Contract, Work From Home
Contract Rate
From £1,000 to £1,200 per day
RoCEv2 networking. Develop infrastructure solutions using NVIDIA Blackwell, B300, GB300 and GB200 platforms. Design both bare-metal and Kubernetes-based GPU environments. Work with Slurm, Kubernetes, NVIDIA GPU Operator, NCCL and GPUDirect. Design high-performance storage solutions for AI workloads. Lead technical discussions with CTOs, AI leaders and infrastructure ...

Platform Architect - Nvidia AI/GB300

Hiring Organisation
Oscar Associates (UK) Limited
Location
London, United Kingdom
Employment Type
Contract
Contract Rate
£700 - £765 per day
data-centre teams. Key Requirements Strong Platform/Infrastructure Architecture experience across compute, storage, networking and Linux. Expert-level Kubernetes architecture experience. Strong Slurm and Run experience - essential. Proven experience with GPU/HPC environments and large-scale AI platforms. Hands-on experience with NVIDIA HGX GB300/NVL72 … including NVLink, NVSwitch and Grace Blackwell architecture. Experience with NVIDIA RTX 6000 series GPU servers. Strong understanding of GPU workload scheduling, partitioning and sharing, including MIG, vGPU and time-slicing Strong understanding of InfiniBand, RoCE, Spectrum-X, GPUDirect RDMA/Storage and high-performance AI fabrics. Experience with Terraform ...