Senior Solutions Architect – Large Scale Neural Networks Inference
Lead technical strategy for AI inference across EMEA by collaborating with AI-native customers and internal teams to solve challenges in latency, efficiency, and scalability. Design and optimize high-performance inference pipelines using NVIDIA's stack, including TensorRT-LLM and Dynamo, and translate real-world deployment insights into product improvements. Work at the intersection of research and engineering to advance large-scale neural network inference in production environments.