Latest Cloud Infrastructure Jobs

Spotlight
Fractile logo

ML Runtime Engineer (Mid-Level and Senior)

This role involves developing and optimizing the runtime stack for AI accelerators, focusing on integrating with open-source ML frameworks like PyTorch and vLLM. The engineer will work closely with hardware and software teams using a co-design approach to enable high-performance inference for large language models. Key responsibilities include building a high-performance runtime in Rust and supporting inference server integrations.

Fractile London, United Kingdom
Hybrid Permanent
NVIDIA logo

Solutions Architect - NVIDIA Cloud Partners and Datacentre Infrastructure

As a Solutions Architect, you will work closely with NVIDIA Cloud Partners and customers to design and implement advanced AI and HPC GPU infrastructure, focusing on datacentre power, cooling, and MEP requirements. You will collaborate with sales teams to secure business opportunities, guide customers through datacentre design, and provide technical expertise to ensure robust and scalable solutions.

NVIDIA Reading, United Kingdom
On-site Permanent
NVIDIA logo

Solutions Architect - NVIDIA Cloud Partners and Datacentre Infrastructure

As a Solutions Architect at NVIDIA, you will focus on designing and implementing advanced AI and HPC GPU infrastructure, working closely with customers to address power, cooling, and MEP requirements. You will collaborate with sales teams, conduct technical meetings, and develop joint solutions to ensure robust and scalable datacentre deployments.

NVIDIA United Arab Emirates
On-site Permanent
OpenAI logo

Software Engineer, Cloud Infrastructure

About the TeamThe Applications Engineering team works across research, engineering, product, and design to bring OpenAI’s technology to consumers and businesses.You’ll join the team responsible for running the core infrastructure that supports products like ChatGPT and the API. The systems...

OpenAI London, United Kingdom
Permanent

Cloud DevOps Engineer AWS

This role involves deploying and maintaining cloud infrastructure for AI and machine learning solutions, building CI/CD pipelines, and supporting the setup of tools for model training and inference. You will work closely with data scientists and engineers to deliver scalable and secure solutions, with a focus on AWS and GenAI projects.

eFinancialCareers London, United Kingdom
Hybrid Permanent
Luminance logo

Senior IT Infrastructure & Network Engineer

This is a fantastic opportunity to join Luminance, the pioneer of Legal-Grade™ AI for enterprise. Backed by internationally renowned VCs and named in both the Forbes AI 50 list of ‘Most Promising Private AI Companies in the World’ and Inc....

Luminance Cambridge, United Kingdom
Hybrid Permanent

Graduate Cloud Software Engineer (2026 start)

As a Graduate Cloud Software Engineer, you will work on diverse projects involving cloud-based systems, AI, and rich UIs. You'll collaborate with multidisciplinary teams to deliver innovative solutions for high-profile clients, often tackling undefined problems and using the latest technologies like AWS, Azure, and Google Cloud.

Cambridge Consultants United Kingdom
On-site Permanent
Wayve logo

Senior Cloud Site Reliability Engineer

As a Senior Cloud Site Reliability Engineer at Wayve, you will build and scale the reliability foundations of their AI cloud platform, focusing on the Model Development Platform and GPU Compute platform. You will define operational standards, improve capacity planning, and ensure resilient and performant cloud infrastructure, working closely with ML and platform teams.

Wayve London, United Kingdom
On-site Permanent
Synthesia logo

Infrastructure Engineer

This role involves maintaining and scaling Kubernetes clusters, managing AWS and GCP cloud environments, and improving CI/CD systems. You will also focus on observability, FinOps practices, and collaborating with product engineers to deploy and monitor production services.

Synthesia London, United Kingdom
Remote Permanent
Wayve logo

Staff Cloud SRE – AI/ML Platform & GPU Compute

This role involves building and scaling the reliability foundations of Wayve's AI cloud platform, including the Model Development Platform and GPU Compute platform. Responsibilities include defining SLOs, improving capacity planning, leading incident response, and designing observability systems to ensure resilient and performant cloud infrastructure.

Wayve London, United Kingdom
On-site Permanent
OpenAI logo

Software Engineer, Infrastructure Reliability

This role involves designing, building, and operating reliable and performant systems that support cutting-edge AI research and global-scale deployments. You will work on improving system resilience, performance, and automation, collaborating closely with cross-functional teams to ensure high reliability and scalability.

OpenAI London, United Kingdom
On-site Permanent

Forward Deployed Engineer, Infrastructure Specialist (Europe/Middle East)

This role involves leading the end-to-end deployment of Cohere's North AI platform in private cloud and on-premises environments. You will partner with enterprise IT teams to assess infrastructure, security, and data management practices, and design deployment strategies that meet client needs while ensuring compliance with data privacy and security standards.

Cohere United Kingdom
Remote Permanent
Amazon logo

Security Engineer, IAM Stores Security

As a Security Engineer, you will design and build security logging pipelines that process billions of events daily, develop monitoring and detection capabilities for AI/ML workloads, and ensure the security of Amazon's global AWS infrastructure. You will also mentor teammates, write production-ready code, and investigate operational issues to prevent recurrence.

Amazon London, United Kingdom
On-site Permanent

AI Platform/ DevOps Engineer

This role involves designing, building, and operating cloud-native AI platform infrastructure across AWS and Databricks. Responsibilities include deploying and scaling containerised services, managing CI/CD pipelines, and implementing observability and security for AI workloads. The position offers deep technical ownership and long-term architectural impact within a growing AI consultancy.

The Portfolio Group Ec4V4Dy, EC4V 4DY, United Kingdom £70,000 – £80,000 pa
On-site Permanent
PolyAI logo

Senior Platform Engineer

This role involves building and maintaining scalable, secure cloud infrastructure on AWS and Azure to support PolyAI's voice assistant platform. The engineer will lead infrastructure automation using Terraform, improve developer experience, and ensure high availability and cost-efficiency across production environments. The position emphasizes technical ownership, continuous improvement, and collaboration with engineering and security teams to evolve cloud architecture.

PolyAI London, United Kingdom
Hybrid Permanent
Ocado logo

Information Security Engineer

As a Security Engineer, you will work within the Cyber Security team to integrate security into CI/CD pipelines, infrastructure-as-code workflows, and developer tooling. Your responsibilities include designing and maintaining security tooling, securing AWS environments, and collaborating with engineering teams to embed security early in the development lifecycle.

Ocado United Kingdom
On-site Contract