Senior Machine Learning Applications and Compiler Engineer, LPX
Develop high-performance compiler and runtime components for NVIDIA's LPX inference stack, optimizing neural network workloads on future spatial processors. Collaborate with hardware teams to co-design features and implement end-to-end optimizations across compiler, runtime, and deployment layers. Focus on graph transformations, scheduling, and memory layout for domain-specific AI hardware.