Be at the heart of actionFly remote-controlled drones into enemy territory to gather vital information.

Apply Now

Reinforcement Learning Engineer

Cubiq Recruitment
London
10 months ago
Applications closed

Related Jobs

View all jobs

Machine Learning Engineer (Reinforcement Learning) London, UK

Machine Learning Engineer

Machine Learning Engineer (LLM) | London, UK

Machine Learning Engineer (LLM)

Machine Learning Quant Engineer - Investment banking/ XVA

Machine Learning Engineer

Reinforcement Learning Engineer (AI/DataOps)

Candidates should take the time to read all the elements of this job advert carefully Please make your application promptly.
Location : Flexible with hybrid options across the UK and Europe
Type : Full-time, Permanent
About the Opportunity
Were partnering with a dynamic, AI/DataOps startup making waves in the AI-driven solutions space, with strong backing from a globally renowned corporation. As a pre-Series A company with operations across Europe, theyre poised for significant growth, tackling industry challenges through cutting-edge computer vision and AI solutions. Theyre on a mission to launch a revolutionary product powered by state-of-the-art computer vision, with an emphasis on RL-based innovations, so they are looking for exceptional Reinforcement Learning Engineers to join their team.
Role Overview
This is an exciting opportunity for a seasoned Reinforcement Learning Engineer with a solid background in computer vision, data management, and AI-driven modelling. This role offers the chance to design and deploy pioneering RL algorithms that are not just theoretical but engineered for real-world applications. Youll take the lead in developing sophisticated reinforcement learning models from scratch, managing extensive data pipelines, and collaborating with a motivated team of data scientists and engineers to deliver impactful solutions.
Key Responsibilities
Algorithm Development : Design and implement advanced RL algorithms, covering model-based and model-free approaches (e.g., Q-learning, DQN, Policy Gradient, Actor-Critic).
Data Management : Oversee data preparation and integrity to support robust model training.
Model Training and Optimisation : Train, fine-tune, and optimise RL models, ensuring scalability and performance.
Simulation and Deployment : Design simulation environments for effective model training and deploy RL models in production for real-world impact.
Collaboration : Work cross-functionally to align AI solutions with business needs and communicate technical results to non-technical stakeholders.
Why Join?
Be part of a high-growth startup backed by one of the worlds leading companies, ensuring stability while driving innovation.
Competitive salary and benefits package.
A supportive, agile work environment where professional growth is actively encouraged.
Flexibility with remote and hybrid work options across the UK and Europe, allowing you to work alongside a passionate team from anywhere.
What Youll Need
Experience : 4+ years in RL development with both model-based and model-free techniques.
Technical Skills : Proficiency in Python, strong understanding of RL concepts, and familiarity with tools like OpenAI Gym, Stable Baselines, and RLLib.
Education : Bachelors in Computer Science or similar, with preference for Masters/Ph.D. in ML or AI.
Ready to be part of a pioneering AI company with global aspirations? Apply today to help shape the future of AI-driven solutions!

TPBN1_UKTJ

Subscribe to Future Tech Insights for the latest jobs & insights, direct to your inbox.

By subscribing, you agree to our privacy policy and terms of service.

Industry Insights

Discover insightful articles, industry insights, expert tips, and curated resources.

AI Hiring Trends 2026: What to Watch Out For (For Job Seekers & Recruiters)

As we head into 2026, the AI hiring market in the UK is going through one of its biggest shake-ups yet. Economic conditions are still tight, some employers are cutting headcount, & AI itself is automating whole chunks of work. At the same time, demand for strong AI talent is still rising, salaries for in-demand skills remain high, & new roles are emerging around AI safety, governance & automation. Whether you are an AI job seeker planning your next move or a recruiter trying to build teams in a volatile market, understanding the key AI hiring trends for 2026 will help you stay ahead. This guide breaks down the most important trends to watch, what they mean in practice, & how to adapt – with practical actions for both candidates & hiring teams.

How to Write an AI CV that Beats ATS (UK examples)

Writing an AI CV for the UK market is about clarity, credibility, and alignment. Recruiters spend seconds scanning the top third of your CV, while Applicant Tracking Systems (ATS) check for relevant skills & recent impact. Your goal is to make both happy without gimmicks: plain structure, sharp evidence, and links that prove you can ship to production. This guide shows you exactly how to do that. You’ll get a clean CV anatomy, a phrase bank for measurable bullets, GitHub & portfolio tips, and three copy-ready UK examples (junior, mid, research). Paste the structure, replace the details, and tailor to each job ad.

AI Recruitment Trends 2025 (UK): What Job Seekers Must Know About Today’s Hiring Process

Summary: UK AI hiring has shifted from titles & puzzle rounds to skills, portfolios, evals, safety, governance & measurable business impact. This guide explains what’s changed, what to expect in interviews, and how to prepare—especially for LLM application, MLOps/platform, data science, AI product & safety roles. Who this is for: AI/ML engineers, LLM engineers, data scientists, MLOps/platform engineers, AI product managers, applied researchers & safety/governance specialists targeting roles in the UK.