Latest Engineering Manager - Core Infrastructur Jobs

W

Machine Learning Engineer, Compiler

This role involves owning the end-to-end ML compilation pipeline for Wayve's autonomous driving technology, working with NVIDIA TensorRT and Qualcomm QNN. Responsibilities include designing and implementing compiler passes, ensuring accuracy and latency, and collaborating with model and training teams.

Wayve United Kingdom

Staff Software Engineer, Node Infra

This role focuses on designing and operating large-scale AI infrastructure, specifically managing the lifecycle of compute nodes across cloud and on-prem environments. The engineer will lead technical strategy for node provisioning, health monitoring, and automated repair systems, ensuring high reliability and efficiency across Anthropic's GPU, TPU, and Trainium fleets. Collaboration with research, inference, and cloud provider teams is central to shaping long-term infrastructure and compute strategy.

Anthropic London, United Kingdom £325,000 – £485,000 pa

Safety Engineer - Free Tier Abuse

Design and build scalable backend systems for detecting and preventing abuse of free-tier services, focusing on AI-driven moderation and guardrails. Develop production-grade APIs, data pipelines, and observability frameworks while collaborating with ML engineers to deploy models at scale. Own end-to-end technical execution of safety infrastructure across a multimodal AI platform.

ElevenLabs United Kingdom
Remote Permanent

PhD Studentship: AI-Driven fault diagnosis for wind turbine generators

This PhD studentship focuses on developing hybrid fault diagnosis methods for wind turbine generators by combining physical drivetrain models with AI-driven data analysis. The research aims to improve early detection of mechanical and electrical faults in offshore wind turbines using advanced condition monitoring systems, with work involving analytical modelling, simulations, finite element analysis, and experimental validation. The project addresses a critical gap in current monitoring systems, which struggle with false alarms and poor detection of electrical faults.

University of East Anglia Norwich, South East England, United Kingdom
Contract
W

Staff Robotics Engineer

Develop and optimise robust, production-grade online calibration and state estimation software for autonomous vehicles, integrating advanced filtering and optimisation algorithms into resource-constrained, hardware-accelerated platforms. Work across the full software lifecycle, from algorithm adaptation and simulation testing to on-vehicle deployment and fleet-wide observability, collaborating with cross-functional teams to ensure reliability, performance, and safety.

Wayve United Kingdom

Software Developer - Biology

This role involves developing scalable software systems to support virtual drug discovery, focusing on integrating bio foundation models into production workflows. The developer will build and optimize pipelines for biological data processing, create tools for data visualization and quality control, and collaborate with ML scientists and engineers. The position sits at the intersection of computational biology and software engineering within a high-growth startup environment.

Helical London, United Kingdom
Hybrid Permanent

AI Accelerator Lead

The AI Accelerator Lead will establish and drive a strategic initiative to unlock enterprise-wide value from AI, combining technical leadership with programme-level influence. Responsibilities include designing and operating the AI Accelerator, exploring advanced AI techniques, and building high-impact prototypes that evolve into scalable solutions.

Key Recruitment Southampton, Hampshire, United Kingdom £70,000 – £92,258 pa
CrowdStrike logo

Sr. SDET - Cloud, Detection Engineer , London)

Design and build scalable cloud-based detection systems that process billions of events daily to identify sophisticated threats across multi-cloud environments. Work closely with security researchers to translate threat intelligence into automated, real-time detection capabilities using distributed systems and data engineering techniques. Develop and optimize components like custom query languages, event orchestration, and correlation engines within a modular, AI-native security platform.

CrowdStrike London, United Kingdom
Hybrid Permanent
NVIDIA logo

Senior Software Engineer, AI Inference Systems

Design and optimize high-performance AI inference systems for large-scale models using cutting-edge GPU hardware. Develop and enhance inference frameworks like vLLM, implement advanced parallelism techniques, and contribute to compiler and kernel optimization. Build scalable scheduling solutions for multi-node, multi-cloud GPU deployments and drive industry benchmarks such as MLPerf.

NVIDIA PLN 292,500 – PLN 650,000 pa

Senior ML Engineer

Design and deploy production-grade NLP systems and large language models tailored for financial services, with a focus on scalability, data governance, and regulated environments. Develop fine-tuned, low-latency AI workflows integrated into cloud infrastructure, while mentoring engineers and advancing evaluation methodologies for LLM outputs. Work across the full ML lifecycle from research to deployment in a remote-first, innovation-driven team.

Aveni United Kingdom
Remote Permanent

Staff+ Software Engineer, Safeguards Infrastructure

Build and maintain foundational systems for AI safety, oversight, and intervention mechanisms, focusing on detecting unwanted model behaviors and preventing misuse. Develop infrastructure for data management, evaluation systems, and tooling for human and agentic review while ensuring high operational reliability at scale. Work closely with cross-functional teams to implement robust, multi-layered defenses in a safety-first AI environment.

Anthropic London, United Kingdom £325,000 – £395,000 pa

Research Engineer, Machine Learning (Reinforcement Learning)

This role involves advancing reinforcement learning in large language models, focusing on developing agentic systems for computer use, code generation, and reasoning. The engineer will build scalable RL infrastructure, design training environments, and collaborate across research and engineering to implement safety and performance improvements at scale.

Anthropic London, United Kingdom £260,000 – £630,000 pa

Research Engineer, Machine Learning (RL Velocity)

This role involves building and optimizing the machine learning infrastructure that supports Anthropic’s reinforcement learning research. You'll work closely with researchers to improve training efficiency, debug performance bottlenecks, and develop tooling that accelerates model development. The position focuses on high-leverage platform improvements that scale across the entire research team.

Anthropic London, United Kingdom £370,000 – £630,000 pa

Staff Infrastructure Engineer, Cluster Infrastructure

This role focuses on designing and managing large-scale compute cluster infrastructure, with a strong emphasis on automation, security, and scalability. The engineer will lead technical strategy for provisioning and lifecycle management of clusters across cloud and on-prem environments, ensuring high-bandwidth interconnectivity and fault tolerance. Responsibilities include cross-team collaboration on compute strategy, operational excellence, and mentoring other engineers.

Anthropic London, United Kingdom £325,000 – £485,000 pa

Staff Software Engineer, Inference

This role involves designing and maintaining large-scale, performance-sensitive distributed systems that serve Claude to millions of users globally. The engineer will work across the full inference stack, building intelligent routing, autoscaling, and deployment systems for AI models running on diverse accelerators across multiple cloud platforms. The position emphasizes real-time system resilience, compute efficiency, and close collaboration with research teams to enable next-generation model development.

Anthropic London, United Kingdom £325,000 – £390,000 pa