Senior Solutions Architect – Large Scale Neural Networks Inference
Lead the technical strategy for large-scale AI inference deployments with EMEA-based AI-native customers, from proof of concept through to production. Architect and optimize high-performance inference pipelines using NVIDIA's stack, including TensorRT-LLM and Dynamo, while translating real-world deployment challenges into product feedback. Collaborate across NVIDIA teams and customer organizations to drive scalable, efficient AI inference solutions addressing latency, cost, and GPU utilization.