AI Infrastructure Engineer, Serving Platform
Design and build scalable, fault-tolerant systems for serving large language models (LLMs) in both research and production environments. Develop internal platforms for LLM capability discovery and integration, with a focus on performance, observability, and system health. Collaborate closely with researchers and engineers to optimize model deployment and support cutting-edge AI applications.