Senior Solutions Architect – Large Scale Neural Networks Inference
This role involves leading the technical strategy for large-scale AI inference deployments across EMEA, working closely with AI-native customers and internal teams at NVIDIA. You will architect high-performance inference pipelines using NVIDIA's stack, optimize for latency and efficiency, and translate real-world deployment challenges into product improvements. The position requires deep expertise in LLM/VLM inference, transformer optimization, and scalable AI infrastructure.