Talent.com
Google
Research Scientist, Reinforcement Learning, DeepMindGoogle • London, England, GB
Research Scientist, Reinforcement Learning, DeepMind

Research Scientist, Reinforcement Learning, DeepMind

Google • London, England, GB
27 days ago
Job type
  • Full-time
Job description

hackajob is collaborating with Google to connect them with exceptional professionals for this role.

As an organization, Google maintains a portfolio of research projects driven by fundamental research, new product innovation, product contribution and infrastructure goals, while providing individuals and teams the freedom to emphasize specific types of work. As a Research Scientist, you'll setup large-scale tests and deploy promising ideas quickly and broadly, managing deadlines and deliverables while applying the latest theories to develop new and improved products, processes, or technologies. From creating experiments and prototyping implementations to designing new architectures, our research scientists work on real-world problems that span the breadth of computer science, such as machine (and deep) learning, data mining, natural language processing, hardware and software performance analysis, improving compilers for mobile platforms, as well as core search and much more.

As a Research Scientist, you'll also actively contribute to the wider research community by sharing and publishing your findings, with ideas inspired by internal projects as well as from collaborations with research programs at partner universities and technical institutes all over the world.

DeepMind’s Reinforcement Learning (RL) team is a long-standing and tight-knit team of collaborative scientists and engineers. We address research issues in reinforcement learning. We design, refine, and scale RL algorithms and deliver meaningful scientific or product impact. Over the past decade, members of the RL team have been instrumental in building DQN, AlphaGo, Rainbow, AlphaZero, MuZero, AlphaStar, AlphaProof and Gemini.

Artificial intelligence will be one of humanity’s most transformative inventions. At Google DeepMind, we are a pioneering AI lab with exceptional interdisciplinary teams focused on advancing AI development to solve complex global challenges and accelerate high-quality product innovation for billions of users. We use our technologies for widespread public benefit and scientific discovery, ensuring safety and ethics are always our highest priority.

We are pushing the boundaries across multiple domains. Our global teams offer diverse learning opportunities and varied career pathways for those driven to achieve exceptional results through collective effort.

Minimum qualifications:

  • PhD in Machine Learning, or equivalent practical experience.
  • 2 years of experience implementing algorithms within research codebases.
  • Experience conducting research in reinforcement learning, including contributions to peer-reviewed publications.
  • Experience designing and executing end-to-end experiments, including setup, analysis, and interpretation.

Preferred qualifications:

  • Experience with advanced reinforcement learning topics, such as RL for sequence models, post-training, preference-based learning, or agentic systems.
  • Familiarity with modern research stacks (e.g., JAX/Flax or PyTorch) and experience scaling experiments.
  • Ability to be comfortable with scaling methodologies, evaluation techniques, and diagnosing complex failure modes.
  • Ability to push projects forward, prioritize effectively, and take initiative.
  • Strong experimental judgment, including selecting appropriate baselines and designing insightful ablations.
  • Excellent communication skills, with a focus on clear presentation of research results.

Responsibilities:

  • Propose and pursue novel research directions by formulating and testing hypotheses.
  • Implement algorithmic ideas and conduct end-to-end experiments, analyzing results and iterating on findings.
  • Design evaluations and ablations to answer key questions and drive research decisions.
  • Build and improve infrastructure to enable research at scale.
  • Communicate research findings through write-ups, presentations, and publications, while fostering a culture of high standards and constructive feedback.
Create a job alert for this search

Research Scientist, Reinforcement Learning, DeepMind • London, England, GB

Similar jobs

Reinforcement Learning Research Scientist Scale Impact

Google DeepMindGreater London, England, GB
Full-time

Google DeepMind is seeking a Research Scientist to set up large-scale tests, deploy promising ideas, and manage deadlines while applying the latest theories to develop new products and technologies... Show more

 • Promoted • New!

Autonomy AI Researcher - Reinforcement Learning

HelsingGreater London, England, GB
Full-time

Helsing is a defence AI company.Our mission is to protect our democracies.We aim to achieve technological leadership, so that open societies can continue to make sovereign decisions and control the... Show more

 • Promoted

Lead ML Scientist - Search & Personalization

MonzoGreater London, England, GB
Full-time

Monzo is seeking a Lead Machine Learning Scientist to drive the development and deployment of models that enhance how customers search for financial products.You will collaborate across Backend, Pr... Show more

 • Promoted

Research Scientist: Visual Generative AI & World Models

graphcoreGreater London, England, GB
Full-time

Graphcore is seeking a Research Scientist to advance AI research at the intersection of visual generative modelling, multimodal learning, world models and hardware-aware machine learning.You will e... Show more

 • Promoted

Research Scientist, Robotics, DeepMind

Software CareersGreater London, England, GB
Full-time

Research Scientist, Robotics, DeepMindat DeepMind • London, UKBack to jobs1.Research Scientist, Robotics, DeepMind—D## Research Scientist, Robotics, DeepMindDeepMind20Research ScientistFull-timeLon... Show more

 • Promoted

Senior Research Scientist - Deep Learning for Markets

P2PGreater London, England, GB
Full-time

Jump Trading Group in Greater London seeks a research scientist to apply deep learning to state-of-the-art problems across quantitative research and monetization.You will lead open-ended projects f... Show more

 • Promoted

Research Scientist (Visual Generative AI & World Models)

CerebrasGreater London, England, GB
Full-time

At Graphcore, we’re building the future of AI compute.We’re a team of semiconductor, software and AI experts, with deep experience in creating the complete AI compute stack – from silicon and softw... Show more

 • Promoted

Lead AI Scientist — LLM Systems & Experimentation

Experimentation JobsGreater London, England, GB
Full-time

Wise is a global technology company building the best way to move and manage the world’s money.The Contact Automation team in London develops an intelligent system to power an automated Wise Assist... Show more

 • Promoted

VP, Frontier ML & World Models

OdysseyGreater London, England, GB
Full-time

Odyssey, a leading AI lab, is hiring a VP of Research to lead the Frontier ML group.You will shape the research agenda across world simulation, multimodality, robotics, and multi-agent systems, and... Show more

 • Promoted

Reinforcement Learning (RL) Engineer, Manipulation - Full UK Visa Sponsorship Available

EasyInfoBlog.com LLCGreater London, England, GB
Permanent

Reinforcement Learning (RL) Engineer, Manipulation.Randstad Technologies Recruitment.We are currently supporting an exciting, top‑notch high‑tech robotics company in London as they assemble an elit... Show more

 • Promoted

Biologics Scientist, ML Platform Lead (Remote, Equity)

BoltzGreater London, England, GB
Remote
Full-time

Boltz is seeking a Principal Scientist in Biologics to shape Boltz Lab’s platform by translating deep biological insight into ML design and evaluation.You will work with ML researchers to define bi... Show more

 • Promoted

Senior Machine Learning Scientist, FinCrime

EngineersOfAIGreater London, England, GB
Full-time

We are reshaping banking by solving real problems for customers and making money work for everyone.Over the past 10 years in the UK, our product offering has expanded from prepaid cards to personal... Show more

 • Promoted

ML Research Scientist Intern – Dynamo Guard / Dynamo Eval / AgentWarden

Dynamo AIGreater London, England, GB
Full-time

At Dynamo AI, an ML Research Scientist Intern will focus on advancing the state of AI evaluation, adversarial robustness, and agent security.You will contribute to novel research in model safety, h... Show more

 • Promoted

Remote ML Researcher – Generative Video & Lip Sync

DNEGGreater London, England, GB
Remote
Full-time

DNEG is seeking a Machine Learning Researcher to advance generative video models with a focus on expression control and lip synchronisation.This role involves collaboration with top researchers to ... Show more

 • Promoted

Applied AI Lead - NLP & LLM Scientist (Real-World ML)

AumniGreater London, England, GB
Full-time

NLP / LLM Scientist - Applied AI ML Lead - Machine Learning Centre of Excellence.Location: London, United Kingdom.The Machine Learning Center of Excellence invites the successful candidate to apply... Show more

 • Promoted

Deep Learning Research Engineer

PlumeraiGreater London, England, GB
Full-time

At Plumerai, we make it easy and affordable for developers to add highly accurate AI to their embedded devices and thereby enable them to create amazing new products.We combine our on-device Tiny A... Show more

 • Promoted

Machine Learning Research Scientist — Drug Discovery Frontiers

Apam 91London, England, GB
Full-time

Apam 91 is looking for a Research Scientist with a focus on machine learning in London.You will be part of a multidisciplinary team aiming to transform drug discovery through innovative computation... Show more

 • Promoted

Research Scientist - World Models

SpAItial AIGreater London, England, GB
Full-time

SpAItial is pioneering the next generation of World Models, pushing the boundaries of generative AI, computer vision, and the simulation of reality.We are moving beyond 2D pixels to build models th... Show more

 • Promoted

Manager, Lead Research Scientist, LLM Agents (Foundational Research)

Thomson ReutersGreater London, England, GB
Full-time

Are you a curious and open‑minded individual with an interest in conducting state‑of‑the‑art foundational machine learning research? Thomson Reuters Labs is seeking Research Scientists to build com... Show more

 • Promoted

Staff Machine Learning Scientist (Recommendations)

Depop LimitedGreater London, England, GB
Full-time

Depop is the community-powered circular fashion marketplace where anyone can buy, sell and discover desirable secondhand fashion.With a community of over 35 million users, Depop is on a mission to ... Show more