Latest infrastructure engineer Jobs

NVIDIA logo

Senior Software Engineer, RL Post-Training Frameworks

This role involves designing and building scalable reinforcement learning (RL) post-training infrastructure that supports the full lifecycle of training, inference, and rollout across heterogeneous hardware. You'll optimize distributed systems for performance, fault tolerance, and efficiency, working closely with AI researchers and contributing to open-source frameworks like VeRL, Miles, and TorchTitan. The role emphasizes deep collaboration across teams to advance AI capabilities through systems innovation.

Remote Permanent
NVIDIA logo

Senior Software Engineer, RL Post-Training Frameworks

This role involves designing and building scalable reinforcement learning post-training infrastructure that operates efficiently from single-GPU experiments to large-scale distributed systems. You'll optimize training-inference-rollout loops across heterogeneous hardware, contribute to open-source RL frameworks like VeRL and TorchTitan, and collaborate with AI researchers and infrastructure teams to improve distributed runtimes such as Ray and Monarch. The position emphasizes fault tolerance, elastic scaling, and integration with next-generation hardware and deep learning tools.

NVIDIA Germany
Remote Permanent
Synthesia logo

ML Platform Engineer

Design and improve platform systems for model training, evaluation, and production serving. Build reliable, scalable infrastructure and tooling for ML workloads, with a focus on automation, observability, and developer experience. Collaborate with researchers and engineers to develop abstractions that reduce operational overhead in GPU and cloud environments.

Synthesia London, United Kingdom
Remote Permanent
Wayve logo

Staff ML Engineer, Gaia

Leads large-scale training and development of video foundation models for autonomous driving, focusing on improving Gaia’s world-model capabilities to generate synthetic driving scenarios. Works closely with research and engineering teams to advance model architecture and training strategies. Operates in a high-impact, fast-paced environment with technical leadership responsibilities.

Wayve London, United Kingdom
Hybrid Permanent
Wayve logo

System Integration Engineer

This role involves integrating and bringing up autonomous driving systems on internal and partner vehicles, working across hardware and software to ensure reliable performance. The engineer will lead system architecture evaluations, develop testing infrastructure, and troubleshoot multidisciplinary issues from prototype to deployment. Emphasis is placed on hands-on integration, embedded systems development, and collaboration across engineering teams to advance Wayve’s AI-driven mobility solutions.

Wayve London, United Kingdom
Hybrid Permanent
Isomorphic Labs logo

ML Research Engineer, London

This role involves developing and optimizing cutting-edge AI models at the intersection of machine learning and drug discovery. You'll work in a collaborative, interdisciplinary environment to translate research into scalable, production-ready systems, focusing on foundational models like Transformers, GNNs, and Diffusion models. The position emphasizes innovation in computational biology and chemistry, with a strong focus on robust codebases, data pipelines, and full-cycle ML development.

Isomorphic Labs London, United Kingdom
On-site Permanent
Synthesia logo

Senior Research Engineer - Audio Post-Training

This role involves advancing high-quality, expressive synthetic voice generation through post-training optimization of AI models. The engineer will work on fine-tuning speech models using techniques like DPO and LoRA, implementing efficiency improvements such as quantization and distillation, and integrating novel architectures like neural codecs and diffusion models. The position is embedded in a research-driven team focused on real-time, production-grade voice synthesis for global enterprise applications.

Synthesia London, United Kingdom
Remote Permanent
Isomorphic Labs logo

Senior Security Engineer (AI Safety), London, Lausanne

This role involves securing AI-driven drug discovery systems by designing risk frameworks, protecting machine learning artifacts, and implementing security controls across AI infrastructure. The engineer will lead incident response for AI-specific threats, enforce compliance with global regulations, and develop safety mechanisms for autonomous agentic workflows and large language models in a high-stakes scientific environment.

Isomorphic Labs United Kingdom
Hybrid Permanent
Entrust logo

Principal Cloud DevOps Engineer

As a Principal Cloud DevOps Engineer, you will manage and improve our existing infrastructure, build the next generation of our platform, and streamline deployment processes. You will work closely with the Security team to address potential threats and provide tools and guidance to engineers and product teams to enhance scalability, stability, and reliability.

Entrust London, United Kingdom
Hybrid Permanent Flexible
Ocado logo

Senior Backend Engineer

This role involves designing and maintaining large-scale, distributed backend systems for Ocado's in-store order fulfilment platform, leveraging Java or Scala in a cloud environment. The engineer will lead technical initiatives, integrate AI tools into development workflows, and contribute to architectural decisions while supporting agile processes and mentoring team members. The position operates in a hybrid model in Sofia, serving a global network of over 1,000 stores.

Ocado United Kingdom
Hybrid Permanent
Entrust logo

Principal Cloud DevOps Engineer

As a Principal Cloud DevOps Engineer, you will manage and improve our existing infrastructure, build the next generation of our platform, and streamline deployment processes. You will work closely with the Security team to address potential threats and provide tools and guidance to engineers and product teams to enhance scalability, stability, and reliability.

Entrust
Hybrid Permanent Flexible

Machine learning Engineer

Design and deploy production-grade machine learning systems for clients across energy, utilities, and environmental sectors. Work hands-on to operationalise models, build scalable ML infrastructure, and translate complex AI concepts for stakeholders. Collaborate with cross-functional teams to deliver real-world impact through responsible AI.

Faculty London, United Kingdom
Hybrid Permanent Clearance Required

AI Implementation Engineer

This role involves hands-on AI solution delivery, from concept to deployment, working closely with IT, operational, and commercial teams. Responsibilities include building secure and scalable AI models, deploying them into production, and ensuring their active use across the business. The role also focuses on rapid prototyping, integration with existing systems, and measuring business impact.

Adria Solutions Manchester, United Kingdom £50,000 – £85,000 pa
Hybrid Permanent

Forward Deploy Engineer

Design and build AI-powered software solutions for government digital transformation, integrating LLMs into production systems and developing scalable applications across the full stack. Work within multidisciplinary teams to deliver end-to-end products from concept to deployment, with a focus on innovation in workforce profiling, job matching, and digital skills platforms.

TXP London, United Kingdom £700 – £800 pd
Hybrid Contract Clearance Required

Research Data Engineer

Design and build scalable data systems to support multi-omics and machine learning workflows in a high-performance research environment. Focus on optimising data pipelines, storage layouts, and access patterns for efficient model training and scientific analysis across heterogeneous datasets.

Relation Therapeutics London, United Kingdom
Hybrid Permanent