11 of 11 Slurm Workload Manager Jobs in the UK

Platform Application Specialist

Hiring Organisation
Appcast
Location
London, UK
AlertManager).Experience with a variety of database platforms (e.g., PostgreSQL, ClickHouse, MSSQL, Redis, FoundationDB).Familiarity with specific middleware (e.g., Kafka, Consul), HPC schedulers (e.g., Slurm), Kubernetes, or workflow orchestrators (e.g., Airflow, Prefect).Experience with on-premise deployments of development tools like GitLab, Artifactory, CICD, or JupyterHub.Knowledge of other programming ...

Scientific Computing Engineer - Vernalis

Hiring Organisation
Vernalis
Location
Cambridgeshire, England, United Kingdom
configurations for the Enterprise IT team. Key Responsibilities HPC Management: Oversee the health and performance of our on-premises HPC cluster ( Bright Computing/SLURM ). Optimise job scheduling and resource allocation across mixed CPU and GPU nodes for computational chemistry and ML applications. DevOps & Container Orchestration: Lead … Linux (RHEL) and SELinux security policies. DevOps: Strong experience with GitLab CI/CD and container orchestration (Docker Swarm). HPC Stack: Proficiency in SLURM workload management and Bright Computing cluster tools. Storage & Identity: Experience with NFS, enterprise storage (EMC2), and managing UID/GID/SID mappings ...

Junior Platform Specialist

Hiring Organisation
Appcast
Location
London, UK
Gitlab, Artifactory or Docker,Experience with infrastructure automation and configuration management, such as Ansible and Terraform,Experience with HPC and orchestration technologies, such as Slurm or Kubernetes,Experience with Databases and Observability systems, such as Elasticsearch, Datadog, Prometheus, PostgreSQL. ...

Enterprise Architect - AI

Hiring Organisation
World Wide Technology
Location
London, UK
Employment Type
Full-time
parallel and high-throughput file systems (e.g. Everpure, WEKA, VAST, NetApp) sized for training and checkpointing workloads.AI Software, MLOps & Generative AIOrchestration & containers: Kubernetes, Docker, Slurm, Run:ai or equivalent GPU scheduling platforms.ML frameworks: PyTorch and TensorFlow at a working, hands-on level.Distributed training: Horovod, DeepSpeed, Megatron-LM, or equivalent … ArgoCD) for continuous, declarative platform delivery.Pipeline orchestration: Kubeflow Pipelines, Apache Airflow, or Argo Workflows to orchestrate multi-stage training, fine-tuning, and inference pipelines.Cluster & workload scheduling: Slurm, Run:ai, and NVIDIA Base Command Manager for GPU job scheduling; Kubernetes-native GPU scheduling including device plugins ...

Senior Solutions Engineer

Hiring Organisation
LJB & Co
Location
City of London, London, United Kingdom
Employment Type
Contract, Work From Home
Contract Rate
From £1,000 to £1,200 per day
RoCEv2 networking. Develop infrastructure solutions using NVIDIA Blackwell, B300, GB300 and GB200 platforms. Design both bare-metal and Kubernetes-based GPU environments. Work with Slurm, Kubernetes, NVIDIA GPU Operator, NCCL and GPUDirect. Design high-performance storage solutions for AI workloads. Lead technical discussions with CTOs, AI leaders and infrastructure ...

Senior Applied Research Engineer - Video Team

Hiring Organisation
Synthesia
Location
London, UK
Employment Type
Full-time
human-centric generationFamiliarity with world/interactive modelsExperience with GANs or VAEsExperience optimizing inference systems for productionOur stackPython, PyTorch, CUDADeepSpeed, distributed training & inferenceSequence parallelismAWS, SLURM, DockerGitHub, CI/CD pipelinesWho you areYou are research-driven but outcome-focusedYou care about shipping, not just publishingYou can explore multiple ideas quickly ...

Principal Cloud Architect - HPC/GPU & AI Platform Solutions

Hiring Organisation
Appcast
Location
London, UK
public cloud platforms. Infrastructure as Code using Terraform and Ansible. Python, Bash, and PowerShell scripting. Kubernetes and container orchestration. HPC cluster management platforms including Slurm, PBS, or Bright Cluster Manager. High-performance networking technologies including RDMA and InfiniBand. MPI and distributed file systems. Cloud-native architectures. AI/… Large Language Models (LLMs) Agentic AI AI Platform Architecture Inference Serving Programming & AutomationPython Bash PowerShell Automation Frameworks Infrastructure Automation HPC TechnologiesSlurm PBS Bright Cluster Manager RDMA InfiniBand MPI Distributed File Systems Customer & ConsultingSolution Architecture Technical Consulting Pre-Sales Executive Presentations Customer Workshops Technical Enablement AI Transformation Strategy Cloud Adoption ...

HPC Senior Technology Consultant

Hiring Organisation
Hewlett Packard Enterprise
Location
Wokingham, Berkshire, UK
Employment Type
Full-time
customer-facing environment.Desirable ExperienceExperience in one or more of the following areas would be advantageous:Cluster management platforms such as Bright Cluster Manager, HPE Performance Cluster Manager (HPCM) or Cray Systems Manager (CSM).SLURM or PBS Pro job schedulers.High-performance networking technologies such as InfiniBand or Slingshot.Parallel ...

Senior AI Platform Engineer

Hiring Organisation
IQVIA
Location
London, UK
Employment Type
Full-time
scalable engineering solutions.Partner with centralised infrastructure teams to design and deliver high-performance compute environments across AWS and on-premises platforms, including GPU infrastructure, Slurm clusters, and migration from ad hoc research workflows.Optimise LLM training and inference workloads, supporting research and product teams in maximising performance, scalability, and reliability … NVIDIA Nsight, DCGM, and related ecosystem technologies.Strong background in AWS cloud services, high-performance computing, distributed systems, containerised environments, and infrastructure automation.Experience with workload orchestration technologies such as Slurm, Kubernetes, Ray, or equivalent distributed compute frameworks.Demonstrated success bridging research and production environments, enabling rapid experimentation while maintaining operational ...

Platform Architect - Nvidia AI/GB300

Hiring Organisation
Oscar Associates (UK) Limited
Location
London, United Kingdom
Employment Type
Contract
Contract Rate
£700 - £765 per day
data-centre teams. Key Requirements Strong Platform/Infrastructure Architecture experience across compute, storage, networking and Linux. Expert-level Kubernetes architecture experience. Strong Slurm and Run experience - essential. Proven experience with GPU/HPC environments and large-scale AI platforms. Hands-on experience with NVIDIA HGX GB300/NVL72 … including NVLink, NVSwitch and Grace Blackwell architecture. Experience with NVIDIA RTX 6000 series GPU servers. Strong understanding of GPU workload scheduling, partitioning and sharing, including MIG, vGPU and time-slicing Strong understanding of InfiniBand, RoCE, Spectrum-X, GPUDirect RDMA/Storage and high-performance AI fabrics. Experience with Terraform ...

Founding GPU Engineer

Hiring Organisation
Fuse Energy Supply
Location
London, UK
Employment Type
Full-time
center systems. You'll work on low-level performance engineering for large-scale compute clusters, helping Fuse build the software layer that ties GPU workload behaviour to energy availability and grid demand.The OpportunityDemand for high-performance compute capacity across the markets we operate in significantly outpaces what … multi-node scaling using NCCL, MPI, or similar communication libraries.Work with data center infrastructure teams on power capping, dynamic voltage/frequency scaling, and workload scheduling strategies that reduce energy cost and carbon intensity.Collaborate with ML/systems engineers to integrate custom kernels into training/inference pipelines.Benchmark against ...