Talent.com
Canva
PhD Research Scientist Intern Reinforcement Learning, ImagesCanva • London, England, UK
PhD Research Scientist Intern Reinforcement Learning, Images

PhD Research Scientist Intern Reinforcement Learning, Images

Canva • London, England, UK
7 days ago
Job type
  • Full-time
Job description

Join the team redefining how the world experiences design.

Hey gday mabuhay kia ora 你好 hallo vítejte!

Were looking for current PhD students ready to bring their research into the real world and help shape the culture of AI at Canva.

Our full-time 16 week AI Research Internship starts in September. During your internship youll work directly with Canvas AI team on a live industry-scale project turning part of your PhD journey into real world impact.

Youll gain hands on experience with real data production infrastructure and real deadlines while learning from and working alongside the researchers and engineerings creating Canvas next generation of AI-powered experiences.

What youd be doing in this role

As Canva scales change continues to be part of our DNA but we like to think thats all part of the fun. This gives you a flavour of the work youd start with and it will likely evolve over time. At the moment this role is focused on:

  • Designing and validating a rubric-guided per-layer VLM judge for RGBA layer decomposition calibrated against human evaluations.

  • Building VLM-based methods for automatic human-aligned evaluation of multi-layer designs.

  • Turning VLM-based evaluators into reward functions to train generative models in a reinforcement learning setting.

  • Distilling those judges into lightweight reward models that score layered images from learned representations at a fraction of the inference cost.

  • Collaborating with research engineering and product teams to move findings toward production and Canvas layered-generation roadmap.

  • Contributing to the broader research community through publication where results support it.

The team builds the groundwork before you arrive baselines reproduced harnesses running data prepared. That means you start on the novel parts in week one rather than spending a month on setup.

Youre probably a match if

  • Youre currently completing a PhD ideally third year or later.

  • A strong diffusion or flow-matching backgroundwith hands-on policy-gradient RL for generative models (GRPO PPO DPO or similar).

  • Experience fine-tuning VLMs (e.g. with LoRA) and designing prompts or rubrics for evaluation tasks.

  • Reward modelling experience preference optimisation pseudo-labelling distillation.

  • You can read a recent paper and reproduce it quickly.

  • You communicate technical work clearly in writing and in presentations.

  • You enjoy working closely with researchers and engineers on hard problems.

  • Juggle several threads at once drop into a new one without losing the last

  • Set your own priorities on a daily basis and between checkpoints

Nice to have

  • PyTorch at scale and the ability to write research code for data processing training and evaluation.

  • Multi-GPU training (FSDP DeepSpeed) and evaluation-harness engineering.

  • Layered or RGBA generation matting or inpainting experience.

  • Familiarity with reward-hacking and score-compression diagnostics or human-evaluation design.

  • Publications or open-source contributions in generative modelling RLHF or multimodal models.

What you should aim to take away

  • Publishable and patentable contributions based on the work youve done

  • A paper draft covering said work with support on publication strategy.

  • Compute base checkpoints preference data and annotation budget provided.

  • Four mentors: a coach for weekly 1:1s plus a specialist lead on each workstream.

  • Work that feeds directly into a product used by hundreds of millions of people.


Additional Information :

Other stuff to know

We make hiring decisions based on your experience skills and passion as well as how you can enhance Canva and our culture. When you apply please tell us the pronouns you use and any reasonable adjustments you may need during the interview process.

We celebrate all types of skills and backgrounds at Canva so even if you dont feel like your skills quite match whats listed above - we still want to hear from you!

Please note that interviews are conducted virtually.


Remote Work :

No


Employment Type :

Intern


Experience: years
Vacancy: 1
Create a job alert for this search

PhD Research Scientist Intern Reinforcement Learning, Images • London, England, UK

Similar jobs

Launchpad VFX Internship – London 2026 (Mentored)

FestybayGreater London, England, GB
Full-time

Framestore is seeking interns for its prestigious Launchpad Internship Program in London for high-end visual effects roles.Applicants should be in their final year of study or recent graduates and ... Show more

 • Promoted

Research Scientist: Visual Generative AI & World Models

graphcoreGreater London, England, GB
Full-time

Graphcore is seeking a Research Scientist to advance AI research at the intersection of visual generative modelling, multimodal learning, world models and hardware-aware machine learning.You will e... Show more

 • Promoted

Forward Deployed AI Scientist, Internship, United Kingdom - BCG X

The Boston Consulting Group GmbHGreater London, England, GB
Internship

Boston Consulting Group partners with leaders in business and society to tackle their most important challenges and capture their greatest opportunities.BCG was the pioneer in business strategy whe... Show more

 • Promoted

Research Scientist, Reinforcement Learning, DeepMind

Google Inc.Greater London, England, GB
Full-time

PhD in Machine Learning, or equivalent practical experience.Experience conducting research in reinforcement learning, including contributions to peer‑reviewed publications.Experience designing and ... Show more

 • Promoted

Campus ML Research Engineer (Intern)

P2PGreater London, England, GB
Full-time

Jump Trading Group is committed to world class research.We empower exceptional talents in Mathematics, Physics, and Computer Science to seek scientific boundaries, push through them, and apply cutt... Show more

 • Promoted

Research Scientist (Visual Generative AI & World Models)

CerebrasGreater London, England, GB
Full-time

At Graphcore, we’re building the future of AI compute.We’re a team of semiconductor, software and AI experts, with deep experience in creating the complete AI compute stack – from silicon and softw... Show more

 • Promoted

Founding Scientist

First Momentum VenturesGreater London, England, GB
Full-time

Deep tech startup unlocking stranded geothermal energy and critical minerals.London (on-site) | Full-time | £50–70k salary + founding equity.Drilling advances are unlocking access worldwide to the ... Show more

 • Promoted

Research Scientist, Reinforcement Learning, DeepMind

Google DeepMindGreater London, England, GB
Full-time

PhD in Machine Learning, or equivalent practical experience.Experience conducting research in reinforcement learning, including contributions to peer-reviewed publications.Experience designing and ... Show more

 • Promoted

NLP & ML Research Internship—Build Real-World AI Solutions

JPMorgan Chase & Co.Greater London, England, GB
Internship

Machine Learning Center of Excellence (MLCOE).Interns will apply machine learning methods to complex domains such as natural language processing and big data analytics.Participants will collaborate... Show more

 • Promoted

Summer Foresight Research & Strategy Intern

Houston ForesightCity Of London, England, GB
Full-time

A leading consulting firm in London seeks an intern for the Foresight team, focusing on sustainable development.The role involves conducting high-quality research and providing analytical support.I... Show more

 • Promoted

Research Scientist Intern (PhD) - Model Team - London

H CompanyGreater London, England, GB
Full-time

PhD - Graduation/Final year Internships available in our.PhD (completed or near completion) in Computer Science, Machine Learning, Mathematics, or a related field.Published in top-tier conferences ... Show more

 • Promoted

Data Scientist Intern, AI Agents Filled Internship · London · ~3 months · £3k–5k monthly

Arct a LGreater London, England, GB
Full-time +1

Arctal builds structured datasets from unstructured financial documents—100,000+ PDFs (fund reports, regulatory filings, investor letters) turned into clean, queryable data that institutional buyer... Show more

 • Promoted

ML Research Scientist Intern – Dynamo Guard / Dynamo Eval / AgentWarden

Dynamo AIGreater London, England, GB
Full-time

At Dynamo AI, an ML Research Scientist Intern will focus on advancing the state of AI evaluation, adversarial robustness, and agent security.You will contribute to novel research in model safety, h... Show more

 • Promoted

AI Research Scientist - Frontier AI Peer Review

OxSciGreater London, England, GB
Full-time

OxSci in London seeks a PhD-level researcher to lead AI-reviewer evaluation, shaping how AI reviewers are judged in science.You will design meta-evaluations, run large-scale expert-annotation studi... Show more

 • Promoted

Data Science Intern — Insights & Dashboards in Music Tech

SpotifyGreater London, England, GB
Full-time

Spotify is seeking Data Science interns in London for a summer internship program starting in June 2024.Interns will conduct data analyses to provide insights that inform business strategies and co... Show more

 • Promoted

PhD Research Scientist Intern - Foundational Research

Thomson ReutersGreater London, England, GB
Full-time

Interested in training and evaluating large‑scale language models (>.B) in a frontier research team focused on AI impact in high‑stakes domains? Thomson Reuters Foundational Research offers an oppo... Show more

 • Promoted

Research Scientist Intern, Grounded Multimodal Understanding (PhD)

Meta Platforms, Inc.Greater London, England, GB
Full-time

Meta is seeking Research Interns to join our Generative AI efforts across modalities (images, video, 3D, audio, etc.We are committed to advancing the field of artificial intelligence by making fund... Show more

 • Promoted

Research Scientist/Engineer — Frontier AI & Evaluation

COL LimitedGreater London, England, GB
Full-time

A leading AI research organization in London seeks Research Scientists and Engineers to develop the 'Science of Scheming.This role involves collaborating with AI developers, studying complex AI dyn... Show more

 • Promoted

AI Ops Intern: Shape Product Strategy & Growth

Duku AIGreater London, England, GB
Full-time

Duku AI, located in Greater London, is seeking an exceptional Account Executive to help define the AI category in software development.You will drive execution across various functions, work closel... Show more

 • Promoted

Research Scientist - World Models

SpAItial AIGreater London, England, GB
Full-time

SpAItial is pioneering the next generation of World Models, pushing the boundaries of generative AI, computer vision, and the simulation of reality.We are moving beyond 2D pixels to build models th... Show more