Latest llm Jobs

NVIDIA logo

Senior MLOps Engineer - DSX Enablement

This role involves building and deploying custom AI solutions on NVIDIA's NeoCloud and partner cloud platforms, with a focus on MLOps pipelines, distributed training, and inference optimization. The engineer will act as a technical advisor to both internal and external customers, diagnosing full-stack AI/ML system issues and contributing to open-source tools and reference architectures. A strong emphasis is placed on performance tuning, scalability, and deep collaboration with infrastructure and framework teams to advance AI capabilities.

NVIDIA
NVIDIA logo

Senior Solutions Architect – Large Scale Neural Networks Inference

Lead technical strategy for AI inference deployments across EMEA, working closely with AI-native enterprises and frontier labs to optimize large-scale neural network inference on NVIDIA's platform. Architect high-performance inference pipelines using frameworks like TensorRT-LLM, vLLM, and SGLang, and translate real-world deployment challenges into product improvements. Focus on solving critical issues around latency, efficiency, memory utilization, and GPU cluster performance.

NVIDIA PLN 292,500 – PLN 650,000 pa
AI Security Institute logo

Research Engineer – Human Influence

The Research Engineer role involves designing and building reinforcement learning environments, leveraging interpretability methods to mitigate concerning model behavior, and developing scalable evaluation pipelines. The work is interdisciplinary, combining AI safety, machine learning, and computational social science.

AI Security Institute London, United Kingdom
Adecco logo

Snowflake Data Architect - London, Wembley

This role involves designing and implementing data architecture using Snowflake, AWS, and AI/ML technologies. Responsibilities include defining data modeling standards, securing data pipelines, and optimizing Snowflake performance. The position also focuses on cross-cloud orchestration, data governance, and preparing the data infrastructure for AI applications.

Adecco London, United Kingdom £80,000 – £90,000 pa
On-site Permanent

Senior AI Engineer - Consulting

The role involves designing and implementing end-to-end AI/ML solutions, building scalable generative AI and agentic AI solutions, and translating business requirements into cloud-native AI architectures. The successful candidate will collaborate with stakeholders to deliver robust, scalable, and responsible AI solutions across major cloud platforms like Azure, AWS, and GCP.

Deerfoot Recruitment Solutions St Paul's, City And County Of the City Of London, EC4M 9AD, United Kingdom £70,000 pa
Hybrid Permanent

Principal Platform Engineer (SRE/Cloud)

This role involves leading the design and evolution of Beamery's cloud platform and infrastructure, with a focus on reliability, scalability, and operational excellence. The Principal Platform Engineer will drive architectural decisions, mentor engineering teams, and ensure robust SRE practices across Kubernetes, observability, incident response, and cost management. A key part of the role is shaping the platform's future in the context of AI-driven talent systems and large-scale data infrastructure.

Beamery London, United Kingdom
Hybrid Permanent

Machine Learning Engineer, Platform

This role involves building and owning end-to-end machine learning components within a generative AI platform, with a focus on retrieval systems, knowledge representation, and RAG pipelines. The engineer will design and implement systems for knowledge retrieval, vector indexing, and context engines that power enterprise AI agents. Work includes developing evaluation frameworks, integrating with enterprise data sources, and collaborating across ML, product, and infrastructure teams to deliver high-impact AI solutions.

Scale AI London, United Kingdom
Hybrid Permanent

Staff ML Engineer | Agentic AI & Applied ML | London | |

Design and implement production-grade AI and ML systems, focusing on Agentic AI, RAG, and LangChain/LangGraph patterns. Establish engineering standards, improve observability and model risk controls, and drive adoption across engineering teams. Work within a regulated environment to enable safe, scalable AI solutions in production.

WeDoTech Bunhill, London, United Kingdom £800 – £1,000 pd
Hybrid Contract

AI Infrastructure Engineer, Sandbox Platform

This role involves building and maintaining a secure, high-performance sandboxing platform for AI agent workflows, focusing on strong isolation, low-latency execution, and developer experience. The engineer will design APIs, optimise resource utilisation, debug production issues, and collaborate with internal teams to shape the platform’s evolution across containerised and virtualised environments. It requires deep systems expertise in Linux internals, virtualisation, and developer-facing tooling.

Scale AI London, United Kingdom

Senior ML Engineer

Design and deploy production-grade NLP systems and large language models tailored for financial services, with a focus on scalability, data governance, and regulated environments. Develop fine-tuned, low-latency AI workflows integrated into cloud infrastructure, while mentoring engineers and advancing evaluation methodologies for LLM outputs. Work across the full ML lifecycle from research to deployment in a remote-first, innovation-driven team.

Aveni United Kingdom
Remote Permanent

Lead Software Engineer

Lead a product squad in technical decision-making and architectural direction, driving AI-powered feature development using LLMs, RAG, and agent frameworks within a regulated financial services environment. Mentor engineers while delivering production-grade, full-stack systems on AWS with Node.js, Python, and React. Collaborate across platform teams to build reusable, compliant-by-design solutions that integrate cutting-edge AI tooling across the SDLC.

Aveni United Kingdom
Remote Permanent

Research Engineer, Machine Learning (Reinforcement Learning)

This role involves advancing reinforcement learning in large language models, focusing on developing agentic systems for computer use, code generation, and reasoning. The engineer will build scalable RL infrastructure, design training environments, and collaborate across research and engineering to implement safety and performance improvements at scale.

Anthropic London, United Kingdom £260,000 – £630,000 pa

Research Engineer, Pretraining Scaling - London

This role involves owning critical parts of Anthropic's production pretraining pipeline for large language models, focusing on reliability, performance optimization, and debugging across hardware, networking, and training dynamics. The engineer will run experiments to improve training efficiency, respond to incidents during model launches, and collaborate closely with research and systems teams. It’s a high-impact, operational role blending deep technical engineering with empirical research in a fast-paced, safety-focused AI environment.

Anthropic London, United Kingdom £260,000 – £630,000 pa

Senior AI Engineer

Design and build production-grade AI applications using Azure, integrating LLMs, RAG, and agentic AI into enterprise systems. Develop full-stack solutions with C#/.NET and React/Angular, focusing on scalable, secure cloud-native architectures. Collaborate with product and business teams to deliver impactful AI capabilities across the organisation.

Sanderson Bristol, United Kingdom £85,000 – £100,000 pa
Hybrid Permanent

Senior AI Engineer

Design and build production-grade AI applications integrating LLMs, RAG, and agentic systems within a financial services environment. Develop full-stack solutions using C#/.NET and React/Angular, with a focus on scalable cloud-native architectures on Azure. Collaborate with product and business teams to deliver secure, reusable AI capabilities that drive measurable business outcomes.

Sanderson London, United Kingdom £85,000 – £100,000 pa
Hybrid Permanent