Latest infrastructure engineer Jobs

NVIDIA logo

Senior Software Engineer, CUDA Core Libraries

Develop and optimize foundational CUDA Core Libraries in C++ and Python, focusing on high-performance parallel algorithms and APIs for GPU computing. Work across the stack to improve developer experience through robust testing, tooling, and documentation. Collaborate with senior engineers and engage directly with users to refine performance and functionality in a large-scale, multi-language codebase.

NVIDIA PLN 292,500 – PLN 650,000 pa
Remote Permanent
NVIDIA logo

Senior Systems Software Engineer, Kubernetes Scale - DGX Cloud

This role involves driving performance and scalability for NVIDIA's DGX Cloud software stack, focusing on Kubernetes and NVIDIA components like GPU Operator, DCGM, and NIM. The engineer will diagnose complex distributed systems issues, build automated testing and monitoring tools, and collaborate with AI teams and open-source communities to optimize large-scale AI infrastructure. Work includes deep performance analysis, CI/CD integration, and contributing to upstream Kubernetes and CNCF projects.

NVIDIA Germany PLN 292,500 – PLN 650,000 pa
Remote Permanent
Entrust logo

Senior Software Test Engineer

Join us at EntrustAt Entrust, we’re shaping the future of identity centric security solutions. From our comprehensive portfolio of solutions to our flexible, global workplace, we empower careers, foster collaboration, and build solutions that help keep the world moving safely.Get...

Entrust London, United Kingdom
Entrust logo

Senior Software Test Engineer

Join us at EntrustAt Entrust, we’re shaping the future of identity centric security solutions. From our comprehensive portfolio of solutions to our flexible, global workplace, we empower careers, foster collaboration, and build solutions that help keep the world moving safely.Get...

Entrust Portugal

Senior Staff+ Software Engineer, Kubernetes Platform

This role involves owning and scaling the Kubernetes control plane for large AI training clusters, including customizing the scheduler for topology-aware ML workloads and ensuring high availability under extreme scale. The engineer will build core platform services like service discovery, develop controllers and operators, and collaborate closely with research and infrastructure teams. The position demands deep expertise in distributed systems, Kubernetes internals, and debugging complex production issues across cloud environments.

Anthropic London, United Kingdom £325,000 – £485,000 pa
On-site Permanent

Senior Full Stack Engineer

Build and own end-to-end customer-facing analytics products using .NET and TypeScript, with full-stack responsibility from frontend to Kubernetes-based backend. Operate production services on a modern cloud-native platform using GitOps, Crossplane, and declarative infrastructure. Focus on scalable APIs, accessible UIs, observability, and secure, cost-aware deployments in a continuous delivery environment.

dunnhumby Manchester, United Kingdom
Hybrid Permanent

Senior ML Systems Engineer, Frameworks & Tooling

Design and maintain core components of a large-scale LLM training framework, focusing on distributed systems, performance optimization, and developer tooling. Work across the ML stack to improve throughput, stability, and reproducibility on multi-node GPU clusters. Build monitoring, debugging, and automation tools to support fast-moving research and production workflows.

Cohere London, United Kingdom
Remote Permanent
PolyAI logo

Forward Deployed AI Engineer (Must be PST timezone)

This role involves deploying and optimizing voice AI solutions for enterprise clients, acting as the technical liaison between PolyAI and customers. The engineer will configure conversational AI systems, integrate with telephony infrastructure like SIP/RTP, troubleshoot production issues, and use data-driven methods to improve system reliability and performance. The position emphasizes hands-on technical work, client collaboration, and rapid problem-solving in real-world AI deployments.

PolyAI Bristol, BS34 5PA, United Kingdom US$150,000 – US$190,000 pa
Remote Permanent
OpenAI logo

Technical Threat Investigator, Threat Intel Engineering - UK

About the TeamSecurity is at the foundation of OpenAI’s mission to ensure that artificial general intelligence benefits all of humanity.The Threat Intelligence team protects OpenAI’s technology, people, research, and infrastructure by proactively identifying and disrupting adversaries who seek to compromise...

OpenAI London, United Kingdom
Remote Permanent
OpenAI logo

Training, Process Management Engineer

About the TeamTraining Runtime designs the core distributed runtime that powers everything from early research experiments to frontier-scale model runs. We work on building robust, scalable, high performance components to support our distributed training workloads. Our priorities are to maximize...

OpenAI London, United Kingdom
Hybrid Permanent
PhysicsX logo

Forward Deployed Software Engineer

About us PhysicsX is a deep-tech company with roots in numerical physics and Formula One, dedicated to accelerating hardware innovation at the speed of software. We are building an AI-driven simulation software stack for engineering and manufacturing across advanced industries....

PhysicsX North Tyneside, NE29 8EP, United Kingdom
Ocado logo

Senior Machine Learning Engineer (E3)

This role involves owning the full machine learning lifecycle for production systems that power intelligent features in e-commerce, such as personalisation, recommendations, and search ranking. The engineer will work within a cross-functional data science team to build scalable, reliable ML systems while improving data quality and promoting best practices in MLOps. A strong focus is placed on technical leadership, system design, and enabling other teams through reusable platform components.

Ocado United Kingdom
Hybrid Permanent

Member of Technical Staff, Training Performance Engineer

This role involves optimizing the performance of large language models during training, focusing on improving throughput and accelerator utilization. You'll design high-performance software, write low-level CUDA and Triton kernels, and develop profiling tools to eliminate bottlenecks. The position sits within a research-driven team working at the intersection of ML and systems engineering, leveraging large-scale infrastructure to advance NLP capabilities.

Cohere London, United Kingdom
Hybrid Permanent

Machine Learning Modeling Lead - DTG Capital Markets

This role involves leading the development and deployment of machine learning models that power the firm's trading systems. Responsibilities include setting the long-term vision for the model portfolio, designing training pipelines, building explainability tools, and mentoring a team of ML researchers. The position requires deep expertise in modern ML methods, real-time systems, and quantitative finance.

eFinancialCareers London, United Kingdom
Remote Permanent

Algorithmic Trading Developer - Berenberg Bank

Design and build high-performance Java-based trading engines and real-time systems for algorithmic trading in equities. Develop algo strategies, liquidity-seeking logic, and real-time trading desktops using React. Work closely with trading desks and engineering teams to enhance electronic trading capabilities within a regulated sell-side environment.

eFinancialCareers London, United Kingdom
On-site Permanent