Senior Machine Learning Applications and Compiler Engineer, LPX
Develop high-performance compiler and runtime components for NVIDIA's LPX inference stack, focusing on optimizing neural network workloads for spatial accelerators. Collaborate with hardware teams to co-design future architectures and implement advanced compilation techniques such as graph transformations and memory optimizations. Engage in research and development of novel approaches to inference on domain-specific processors, with opportunities to publish and present findings.