Latest llm Jobs

NVIDIA logo

Senior Solutions Architect – Large Scale Neural Networks Inference

Lead technical strategy for AI inference across EMEA by collaborating with AI-native customers and internal teams to solve challenges in latency, efficiency, and scalability. Design and optimize high-performance inference pipelines using NVIDIA's stack, including TensorRT-LLM and Dynamo, and translate real-world deployment insights into product improvements. Work at the intersection of research and engineering to advance large-scale neural network inference in production environments.

NVIDIA France
NVIDIA logo

Senior C++ Software Engineer – AI Developer Tools

You will architect and develop production-quality C++ components and integrations for NVIDIA’s AI-powered developer-tools ecosystem, focusing on performance-sensitive systems and agentic workflows. You will work closely with multidisciplinary teams to drive technical initiatives from early exploration to production deployment.

NVIDIA United Kingdom

Applied AI Engineer, Beneficial Deployments (Life Sciences)

This role involves serving as the primary technical advisor to life sciences research institutions, helping them adopt and integrate Claude into their scientific workflows. The Applied AI Architect will translate complex research needs into scalable AI solutions, lead technical enablements, and design cohort-based accelerators to drive broad impact across academia and mission-driven organizations. By bridging technical depth with relationship-building, the role focuses on making AI a seamless part of discovery and development in life sciences.

Anthropic London, United Kingdom £165,000 – £190,000 pa

Manager, Technical Deployment

This role involves managing a team of Technical Deployment Leads in the Financial Services sector, ensuring successful AI project delivery, and building strong relationships with executive stakeholders. Responsibilities include hiring, setting delivery standards, and leading complex engagements.

Anthropic London, United Kingdom

Applied AI Engineer, Agents & Automations

This role involves building and improving AI-powered product experiences within Cohere's North platform, focusing on creating reliable and user-friendly AI agents and workflows. Responsibilities include designing new interfaces, building evaluation systems, and working closely with product, design, and customer teams to enhance AI reliability and user trust.

Cohere United Kingdom
Remote Permanent

Technical Program Manager, AI Delivery for Public Sector & Defence, UK

This role involves leading end-to-end technical program delivery for AI deployments within UK public sector and defence organisations. The individual will act as a bridge between Cohere’s advanced AI capabilities and government stakeholders, managing complex procurement, security, compliance, and integration requirements. Success requires navigating air-gapped environments, high-stakes governance, and multi-team coordination to ensure responsible, secure, and scalable AI adoption.

Cohere London, United Kingdom
Hybrid Permanent Clearance Required
NVIDIA logo

Senior C++ Software Engineer – AI Developer Tools

You will architect and develop production-quality C++ components and integrations for NVIDIA’s AI-powered developer-tools ecosystem, focusing on performance, reliability, and maintainable interfaces. You will also contribute to agentic workflows, retrieval systems, and enterprise integrations, working closely with multidisciplinary teams.

NVIDIA
Remote Permanent
NVIDIA logo

Tech Engagement Lead, AI Labs - EMEA

This role involves leading technical partnerships with frontier AI labs across EMEA to integrate NVIDIA's full stack—from GPUs to software libraries—into advanced AI workloads including training, inference, and emerging domains like agents and robotics. The individual will diagnose performance bottlenecks, drive platform optimisations, and translate research trends into product roadmap influence. It’s a strategic position bridging deep technical engagement with product strategy in the fast-evolving generative AI landscape.

NVIDIA France
NVIDIA logo

Tech Engagement Lead, AI Labs - EMEA

This role involves building deep technical partnerships with leading AI research labs and model builders to integrate NVIDIA's GPU hardware, systems, and software stack into their AI development workflows. The candidate will identify performance bottlenecks, drive platform optimisations across training, post-training, and inference, and translate technical insights into product roadmap influence. They will also help shape collaboration strategies, support technical assessments of emerging AI labs, and enable public showcases of joint innovation at events like GTC.

NVIDIA
NVIDIA logo

Senior C++ Software Engineer – AI Developer Tools

You will architect and develop production-quality C++ components and integrations for NVIDIA’s AI-powered developer-tools ecosystem, focusing on Genie, a company-wide AI knowledge and developer-productivity service. Your role involves building reliable, performance-sensitive systems, designing maintainable interfaces, and contributing to agentic workflows and retrieval systems.

NVIDIA
NVIDIA logo

Senior Software Engineer, AI Inference Systems

This role involves building and optimizing high-performance AI inference systems for large-scale models on NVIDIA GPUs. You'll contribute to open-source frameworks like vLLM, optimize GPU kernels and compilers, design benchmarking methodologies, and work on scheduling for multi-node, multi-cloud deployments. The position blends deep systems engineering, performance optimization, and research to advance the state of accelerated AI computing.

NVIDIA Germany PLN 292,500 – PLN 650,000 pa
Remote Permanent
NVIDIA logo

Senior Software Engineer, AI Inference Systems

Design and optimize high-performance AI inference systems for large-scale models, focusing on GPU kernel development, compiler optimization, and distributed inference frameworks. Work on vLLM, speculative decoding, and MLPerf benchmarking while collaborating across compiler, scheduling, and performance teams. Contribute to open-source and publish research to advance ML systems.

NVIDIA PLN 292,500 – PLN 650,000 pa
Remote Permanent
NVIDIA logo

Senior Software Engineer, AI Inference Systems

This role involves building and optimizing AI inference systems for large-scale models using NVIDIA's latest GPU hardware. You'll work on high-performance inference stacks, GPU kernel optimization, compiler infrastructure, and benchmarking, while collaborating with cross-functional teams. The position emphasizes research contributions, open-source development, and deployment across multi-GPU and cloud environments.

NVIDIA PLN 292,500 – PLN 650,000 pa
Remote Permanent
NVIDIA logo

Senior Software Engineer, AI Inference Systems

Design and optimize high-performance AI inference systems for large-scale models on NVIDIA GPUs, focusing on frameworks like vLLM, kernel optimization, compiler infrastructure, and distributed execution. Develop benchmarking methodologies, contribute to MLPerf, and integrate cutting-edge research into production software. Work across compiler, scheduling, and performance teams to push the limits of accelerated computing in multi-GPU and multi-cloud environments.

NVIDIA PLN 292,500 – PLN 650,000 pa
Remote Permanent
NVIDIA logo

Senior Software Engineer, AI Inference Systems

This role involves building and optimizing AI inference systems for large-scale models, focusing on high-performance GPU stacks, kernel optimization, and multi-node deployment. The engineer will contribute to frameworks like vLLM, develop compiler infrastructure, lead benchmarking efforts including MLPerf, and integrate cutting-edge research into production software. Collaboration across compiler, scheduling, and performance teams is central to advancing NVIDIA’s accelerated computing platform.

NVIDIA PLN 292,500 – PLN 650,000 pa
Remote Permanent