Latest infrastructure engineer Jobs

Spotlight
IC Resources logo

ML Systems Engineer

Develop and maintain a software 'digital twin' to model and benchmark a high-speed knowledge layer for agentic AI systems. Design and run experiments to evaluate performance improvements from semantic memory, using GraphRAG and enterprise knowledge graphs, while aligning software models with hardware development. Produce rigorous benchmarks and evidence for technical validation, fundraising, and pilot support.

IC Resources Edinburgh, United Kingdom
On-site Permanent

Staff Infrastructure Engineer, Cluster Infrastructure

This role involves leading the technical strategy for agent-driven automation in cluster lifecycle management, ensuring secure, scalable, and fault-tolerant compute infrastructure across cloud and on-prem environments. You'll collaborate with research, product, and security teams to shape long-term infrastructure direction, with a focus on high-bandwidth interconnectivity and operational excellence. The position emphasizes mentorship, cross-team alignment, and driving innovation in large-scale cluster provisioning and management.

Anthropic London, United Kingdom £325,000 – £485,000 pa
On-site Permanent

Machine Learning Systems / AI Infrastructure Engineer- Quant / Systematic Trading Firms

Engineer cutting-edge machine learning infrastructure at scale within high-performance quantitative trading environments. Focus spans distributed training, GPU optimisation, low-latency inference, and full-stack ML systems, integrating hardware and software to accelerate research and production deployment. Work closely with researchers to build robust platforms for rapid model iteration and deployment on massive compute estates.

eFinancialCareers London, United Kingdom £250,000 – £700,000 pa
OpenAI logo

Software Engineer, Model Deployment- ChatGPT Engineering

Design and operate software managing large-scale GPU clusters for ChatGPT inference, building automation and observability systems to improve fleet efficiency, reliability, and developer productivity. Work across infrastructure, research, and product teams to optimize compute utilization and build AI-powered operational tooling.

OpenAI London, United Kingdom
Hybrid Permanent
OpenAI logo

Software Engineer, Codex Core Agents

About the TeamThe Codex Core Agent team builds the kernel of Codex. We own making the agent better, accelerating research, and making those improvements real in production for our users.That means working across the systems that make Codex actually function...

OpenAI London, United Kingdom
Permanent

Senior ML Ops Engineer

As a Senior MLOps Engineer, you will design and maintain scalable MLOps infrastructure, deploy and manage machine learning models using Azure ML and AKS, and build CI/CD pipelines. You will also implement observability and monitoring frameworks, and collaborate with engineering and delivery teams to support AI solutions.

Harnham - Data and Analytics Recruitment London, United Kingdom £75,000 – £85,000 pa

Senior ML Engineer

Senior ML Ops EngineerLondon (Hybrid, 1-2 days per week) | £75,000 - £85,000 + 10% BonusThis is an opportunity to join a growing AI and data organisation that is using advanced analytics and machine learning to help public sector organisations...

Harnham - Data and Analytics Recruitment London, United Kingdom £75,000 – £85,000 pa
NVIDIA logo

Senior Solutions Architect – Large Scale AI Training

This role involves working closely with leading AI research institutions and model builders across EMEA to solve complex challenges in large-scale distributed training and alignment of neural networks. The Senior Solutions Architect will guide customers on optimizing training efficiency, infrastructure, and software stacks using NVIDIA's full AI training platform, including tools like Megatron-LM, NeMo, and Nemotron. The position bridges technical expertise, customer engagement, and product feedback, with a strong focus on cutting-edge AI workloads such as Mixture-of-Experts models and reinforcement learning-based fine-tuning.

NVIDIA PLN 292,500 – PLN 507,000 pa
NVIDIA logo

Senior Solutions Architect – Large Scale AI Training

This role involves working closely with leading AI research institutions and model builders across EMEA to solve complex challenges in large-scale distributed training and alignment of neural networks. You will guide customers on optimizing infrastructure and software stacks for training foundation models, including MoE and reinforcement learning workflows, while influencing NVIDIA's product roadmap. The position bridges cutting-edge research and engineering, with a focus on real-world deployment of advanced AI training techniques using frameworks like PyTorch, Megatron-LM, and NeMo.

NVIDIA France PLN 292,500 – PLN 507,000 pa

Associate Director, Data Engineering

Lead and scale a unified data platform that integrates wet and dry lab data to accelerate drug discovery. You'll build and manage a high-performing data engineering team, ensure FAIR data principles, and drive scalable, secure infrastructure using modern orchestration and cloud technologies. The role bridges computational science, engineering, and lab operations in a highly interdisciplinary environment.

Relation Therapeutics London, United Kingdom
On-site Permanent
Isomorphic Labs logo

Software Engineer (ML Infrastructure), London

This role involves building and operating a scalable inference platform to serve cutting-edge machine learning models for scientific drug discovery applications. The engineer will focus on distributed systems, Kubernetes-based infrastructure, and production reliability, with responsibilities spanning development, CI/CD, observability, and user support. The position operates within an interdisciplinary AI-driven research environment, requiring deep technical ownership and first-principles thinking.

Isomorphic Labs London, United Kingdom
On-site Permanent

Research Engineer - Data Infrastructure

This role involves building and maintaining large-scale data infrastructure to support cutting-edge AI models, with a focus on data pipelines, curation strategies, and tooling for processing massive datasets. The engineer will develop classifiers and quality filters, design deduplication and augmentation systems, and enable efficient data exploration. The work directly impacts model performance through high-quality, scalable data engineering and infrastructure innovation.

ElevenLabs United Kingdom
Remote Permanent
Isomorphic Labs logo

Staff Software Engineer (Inference Platform), London

This role involves leading the reliability and scalability of AI/ML infrastructure for drug discovery, focusing on GPU/TPU systems, Kubernetes orchestration, and inference services. The engineer will design test harnesses, improve monitoring and CI/CD stability, and ensure high-throughput model serving. Work is closely aligned with research and applied ML teams to support biotech R&D at digital speed.

Isomorphic Labs London, United Kingdom
On-site Permanent

Associate Director, Platform Engineering

Lead the platform engineering function for a TechBio company pioneering drug discovery through single-cell multi-omics and machine learning. Define and evolve the hybrid on-premises and cloud infrastructure strategy, with a focus on Kubernetes, MLOps, scientific compute environments, and platform security. Develop the team and platform to support data-intensive genomic pipelines, model training, and CI/CD maturity while ensuring robust observability, disaster recovery, and compliance with clinical-stage data governance standards.

Relation Therapeutics London, United Kingdom
Permanent

Member of Technical Staff, Integration/RL Team (Research Engineer)

This role involves developing and scaling machine learning algorithms and infrastructure for large-scale, distributed reinforcement learning. You will design and implement high-performance software, optimize post-training algorithms, and collaborate with engineering and scientific teams to enhance the quality of the post-training codebase.

Cohere London, United Kingdom
Remote Permanent

Performance Engineer (Junior) | AI Infrastructure | Cambridge

This role involves working alongside senior engineers to build performance models and calculators for AI infrastructure, predicting the efficiency and cost-effectiveness of different setups. You'll work with real metrics from live training and inference jobs, using your strong background in computer architecture and experience with GPU code and profiling tools.

Pure Resourcing Solutions Dry Drayton, Cambridgeshire, United Kingdom £55,000 – £70,000 pa
Hybrid