Latest Cloud Infrastructure Jobs

Software Engineer (Safety)

As a Software Engineer in Faculty's National Security and AI Safety team, you will build and deploy production-grade machine learning systems, collaborate with cross-functional teams, and ensure the technical feasibility and impact of AI solutions. You will work on high-stakes, high-impact projects, focusing on AI safety and responsible development.

Faculty London, United Kingdom
Permanent

Machine Learning Engineer

Design and deploy production-grade machine learning systems for high-impact clients in national security and AI safety. Work across the full ML lifecycle, from prototyping to scalable deployment, using cloud infrastructure and modern MLOps practices. Collaborate with cross-functional teams to deliver secure, reliable AI solutions that solve real-world challenges.

Faculty London, United Kingdom
Hybrid Permanent
OpenAI logo

Software Engineer, Agent Infrastructure

About the TeamThe Agent Infrastructure team at OpenAI is responsible for building systems that enable training and deployment of highly useful AI agents, both internally and for the world.We work hand-in-hand with researchers to design and scale the environment in...

OpenAI London, United Kingdom
Permanent
NVIDIA logo

Senior HPC AI Cluster Engineer

This role involves designing, implementing, and maintaining large-scale HPC/AI clusters, managing job/workload schedules, and developing CI/CD pipelines. You will work closely with HPC, OS, GPU compute, and systems specialists to architect and optimize performance platforms, and support R&D activities.

NVIDIA

Graduate AI Architect

This role involves supporting the design and development of AI solutions within a telecommunications environment, with a focus on AI integration, network optimisation, and cloud-based data projects. The candidate will work alongside technical teams to implement responsible AI practices and contribute to real-world AI initiatives. A structured training programme and mentorship support professional growth in AI and cloud technologies.

Global Tech Recruitment London, United Kingdom £42,000 – £45,000 pa

Data Engineer

This role involves designing and owning cloud-native data platforms for defence, government, and national security clients. You'll work with big, messy, multi-source datasets, build data lakes and pipelines on AWS or Azure, and collaborate closely with data scientists, software engineers, and architects.

Forward Role Cheltenham, United Kingdom £50,000 – £85,000 pa

MLOps Engineer

This role involves owning and developing the MLOps infrastructure for a growing AI platform, building and maintaining orchestration pipelines, deploying and managing ML workloads on Kubernetes and cloud platforms, and working closely with data scientists and AI teams to scale deployments.

Harnham - Data and Analytics Recruitment London, United Kingdom £75,000 – £85,000 pa

ML Data & Platform Engineer

This role involves building and maintaining scalable data pipelines and ML infrastructure to support speech AI models. You'll work across the full lifecycle—from data acquisition and preparation to model training, evaluation, and production serving—while improving data quality, observability, and MLOps practices. The position emphasizes end-to-end ownership of distributed systems and close collaboration with ML teams to accelerate model deployment.

Speechmatics London, United Kingdom

ML Data & Platform Engineer

This role involves building and maintaining scalable data pipelines and ML infrastructure for training and serving speech AI models. You'll work across the full stack to improve data acquisition, model deployment, and system reliability, with a focus on solving data challenges unique to speech AI. The position emphasizes ownership, end-to-end problem-solving, and close collaboration with the ML team to accelerate model iteration and productionisation.

Speechmatics Cambridge, United Kingdom

Lead AI Engineer

This role involves designing and implementing AI solutions to improve operational efficiency, defining LLM architecture, and building AI tools like assistants and workflow automation. You will work within an Azure cloud environment, applying best practices in MLOps and DevOps, and collaborate closely with product teams and senior leadership.

Harnham - Data and Analytics Recruitment London, United Kingdom £85,000 – £95,000 pa
Hybrid Permanent
NVIDIA logo

Senior MLOps Engineer - DSX Enablement

Develop and optimize full-stack AI/ML systems for internal and external customers, focusing on MLOps pipelines, distributed training, and inference performance. Build open-source tools and reference architectures to scale AI workloads on NVIDIA platforms and cloud partners. Act as a technical advisor for complex production issues across hardware, software, and infrastructure layers.

NVIDIA £292,500 – £650,000 pa
Remote Permanent

MLOps Engineer

Design and operate a secure, scalable AI platform infrastructure, building CI/CD pipelines and deployment workflows in collaboration with data science and engineering teams. Implement observability, automate infrastructure using IaC, and ensure high availability and compliance. Focus on performance, cost efficiency, and seamless integration of AI workloads and agents.

DGH Recruitment London, City And County Of the City Of London, United Kingdom £80,000 – £100,000 pa

Senior AI Engineer

This role involves designing, building, and scaling enterprise-grade AI platforms, transforming AI research into production systems. You'll work on complex AI services, data pipelines, and cloud-native architectures, collaborating with Data Scientists and Software Engineers to productionize ML and NLP models.

Infused Solutions London, United Kingdom £65,000 – £75,000 pa
Hybrid Permanent Flexible

Staff Software Engineer, Kubernetes Platform

This role involves owning and scaling the Kubernetes control plane for large AI training clusters, including customizing the scheduler for topology-aware ML workloads and ensuring high availability under extreme scale. The engineer will build core platform services like service discovery, develop controllers and operators, and collaborate closely with research and infrastructure teams. The position demands deep expertise in distributed systems, Kubernetes internals, and debugging complex production issues across cloud environments.

Anthropic London, United Kingdom £325,000 – £485,000 pa
On-site Permanent

Platform Engineer

This role involves designing, building, and enhancing a secure and scalable Databricks data platform, developing data pipelines, and contributing to platform architecture and best practices. You will work closely with Product, Analytics, and Data Science teams to deliver effective data solutions and support platform operations.

Gattaca Surrey, United Kingdom £75,000 – £85,000 pa