Senior Machine Learning Applications and Compiler Engineer, LPX
Develop high-performance compiler and runtime components for NVIDIA's LPX inference stack, focusing on end-to-end optimization of neural network workloads. Collaborate with hardware teams to co-design future architectures and implement novel compilation techniques for spatial accelerators. Profile and benchmark performance while contributing to libraries and tools for cross-platform model deployment.