Talent.com
G-Research
NLP Performance EngineerG-Research • London, England, UK
NLP Performance Engineer

NLP Performance Engineer

G-Research • London, England, UK
9 days ago
Job type
  • Full-time
Job description

We tackle the most complex problems in quantitative finance by bringing scientific clarity to financial complexity.

From our London HQ we unite world-class researchers and engineers in an environment that values deep exploration and methodical execution because the best ideas take time to evolve. Together were building a world-class platform to amplify our teams most powerful ideas.

As part of our engineering team youll shape the platforms and tools that drive high-impact research designing systems that scale accelerate discovery and support innovation across the firm.

Take the next step in your career.

The role

G-Research is investing in how we apply large language models (LLMs) and other Natural Language Processing (NLP) techniques across the firm. We are looking for an exceptional NLP Performance Engineer to join our NLP Engineering team and take ownership of large-scale LLM inference performance.

This is a specialist Quantitative Developer role. Like all our Quantitative Developers you will work alongside researchers to bring their ideas to life with a particular focus on maximising the performance of LLMs.

This is a hands-on high-impact role. You will design and implement techniques that improve the performance cost-efficiency and capabilities of inference workloads on cutting-edge compute infrastructure enabling researchers and engineers to make the best use of current and future systems.

Working closely with research teams and infrastructure engineers you will profile and analyse workloads eliminate bottlenecks and develop reference solutions. Your work will help shape the tooling and infrastructure that underpins our NLP capabilities.

Key responsibilities of the role include:

  • Profiling benchmarking and optimising large-scale LLM inference workloads across our compute infrastructure
  • Ensuring efficient deployment of the latest models across a range of GPU architectures adapting the inference stack as hardware evolves
  • Designing and implementing inference optimisations while maintaining output quality
  • Developing reference implementations libraries and tooling to improve the efficiency and reliability of NLP workloads
  • Collaborating with researchers senior stakeholders and engineers to design optimised solutions
  • Working with systems architecture and platform teams to evolve the compute stack and influence long-term platform decisions

Who are we looking for

Were looking for an engineer who combines deep knowledge of LLM inference with strong software engineering skills and a scientific approach to performance.

The ideal candidate will have the following skills and experience:

  • A Bachelors Masters or PhD in computer science or equivalent experience
  • Proven experience profiling benchmarking and optimising large-scale LLM inference workloads
  • A scientific evidence-led approach to performance optimisation using rigorous benchmarking and reproducible measurement
  • Deep understanding of transformer inference including prefill versus decode KV-cache behaviour attention variants and performance bottlenecks
  • Hands-on experience with LLM serving frameworks such as vLLM SGLang TensorRT-LLM or TGI and the PyTorch ecosystem
  • Experience with inference optimisation techniques including quantisation speculative decoding and model parallelism across modern GPU architectures
  • Strong software engineering skills including Python CUDA and building reliable systems for machine learning workloads
  • Strong communication skills with the ability to collaborate across research infrastructure and engineering teams

Why join us

  • Highly competitive compensation plus annual discretionary bonus
  • Lunch provided (viaJust Eat for Business) and dedicated barista bar
  • 35 days annual leave
  • 9% company pension contributions
  • Informal dress code and excellent work/life balance
  • Comprehensive healthcare and life assurance
  • Cycle-to-work scheme
  • Monthly company events

G-Research is committed to cultivating and preserving an inclusive work environment. We are an ideas-driven business and we place great value on diversity of experience and opinions.

We want to ensure that applicants receive a recruitment experience that enables them to perform at their best. If you have a disability or special need that requires accommodation please let us know in the relevant section


Required Experience:

IC


Employment Type : Full-Time
Experience: years
Vacancy: 1
Create a job alert for this search

NLP Performance Engineer • London, England, UK

Similar jobs

NLP/LLM AI Engineer — Cloud & MLOps

InformaCity Of London, England, GB
Full-time

A leading global information services company is seeking an AI Engineer to design, develop, and deploy advanced AI solutions focusing on Natural Language Processing and Large Language Models.Ideal ... Show more

 • Promoted

Refineries Data & LP Modeling Specialist

MediumGreater London, England, GB
Full-time

Kpler’s Oil & Chemicals business unit is rapidly advancing its real-time analytics and is looking for a technically sharp, data-literate Refineries Analyst to join its growing global team.In this r... Show more

 • Promoted

Lead Data Engineer - Medical NLP & Analytics (Remote)

MedShrGreater London, England, GB
Remote
Full-time

MedShr in Greater London seeks an engineer to join their team, focusing on developing data services to extract insights from clinical data.The role involves architecting data flows and implementing... Show more

 • Promoted

NLP & ML Engineer for Real-Time Data Enrichment

MeltwaterGreater London, England, GB
Full-time

Meltwater in London is seeking a passionate AI Engineer specializing in NLP and ML to contribute to Data Enrichment efforts within Meltwater.You will design, test, and implement machine learning so... Show more

 • Promoted

AI Data Engineer: NLP & LLM ETL in Snowflake/Azure

Medpace, Inc.City Of London, England, GB
Full-time

Data Engineer to join our AI team.This position involves working collaboratively to manage unstructured data and contribute to projects crucial to our success.Ideal candidates will hold a Bachelor'... Show more

 • Promoted

Senior AI Engineer — Remote/Hybrid LLM & NLP Leader

BluefishGreater London, England, GB
Remote
Full-time

Bluefish is seeking a Senior AI Engineer to lead the development of AI-driven products in London.The ideal candidate will design and enhance LLM technologies, ensuring high performance while collab... Show more

 • Promoted

Applied AI ML Director - NLP / LLM and Graphs

J.P. MorganLondon, England, GB
Full-time

The Chief Data & Analytics Office (CDAO) at JPMorgan Chase is responsible for accelerating the firm's data and analytics journey.This includes ensuring the quality, integrity, and security of the c... Show more

 • Promoted

NLP ML Engineer for InsurTech — London (Remote Flex)

NLP PEOPLEGreater London, England, United Kingdom
Remote
Full-time

NLP PEOPLE is looking for a Machine Learning Engineer to join their team in Central London.This full-time role offers a competitive salary of up to £165,000, with stock options and benefits.The com... Show more

 • Promoted

Remote Performance Engineer: ML Training & Kernels

CohereGreater London, England, GB
Remote
Full-time

Cohere is seeking a Performance Engineer in Greater London to optimize the performance of advanced language models and systems.You will focus on improving model training metrics and ensuring high a... Show more

 • Promoted

Senior Applied AI Lead Engineer - NLP/LLM & Graphs

J.P. MORGANGreater London, England, United Kingdom
Full-time

MORGAN is seeking an NLP / LLM Scientist - Applied AI ML Lead to advance machine learning at the Center of Excellence.The role emphasizes collaboration with business and tech teams to deploy produc... Show more

 • Promoted

Machine Learning Performance Engineer

XTX MarketsGreater London, England, GB
Full-time

XTX Markets is a leading algorithmic trading firm which uses state-of-the-art machine learning technology to produce price forecasts for over 50,000 financial instruments across equities, fixed inc... Show more

 • Promoted

Senior Endpoint DLP Engineer (Kernel/Go/C)

FortinetGreater London, England, United Kingdom
Full-time

A leading cybersecurity firm based in London is seeking a Senior Software Engineer to develop and enhance the FortiDLP endpoint agent.The role involves implementing new features, maintaining curren... Show more

 • Promoted

NLP/LLM Data Engineer (Contract)

Future Talent GroupLondon, England, GB
Full-time

A leading recruitment agency in London is seeking a highly skilled Contract Data Engineer to design and optimize data pipelines for NLP and LLM applications.Ideal candidates should have expertise i... Show more

 • Promoted

AI Data Engineer: NLP & ETL for Unstructured Data

MedpaceGreater London, England, GB
Full-time

Medpace is seeking a full-time, office-based Data Engineer to join our AI team in Greater London.In this role, you will work with various data handling projects, from unstructured content extractio... Show more

 • Promoted

NLP Engineer - Text Analytics & LLM (Hybrid, London)

Xantura LimitedGreater London, England, GB
Full-time

Xantura Limited is looking for an AI NLP Engineer to enhance its text analytics platform in London.You will own the development of the NLP system that extracts structured intelligence from unstruct... Show more

 • Promoted

Senior NLP Researcher: LLMs, Conversational AI & Translation

Lightspeed StudiosGreater London, England, GB
Full-time

Lightspeed Studios in Greater London is looking for a candidate focused on NLP research and development.The successful applicant will take responsibility for optimizing NLP algorithms and exploring... Show more

 • Promoted

Machine Learning Performance Engineer

Quant Blueprint LLCGreater London, England, GB
Full-time

We are looking for an engineer with experience in low-level systems programming and optimisation to join our growing ML team.Machine learning is a critical pillar of Jane Street's global business.O... Show more

 • Promoted

Neural Network Performance Engineer

ThehumanoidGreater London, England, GB
Full-time

Here at Humanoid, we believe in a future where robots amplify human potential.That’s why we’ve set out on a mission to build the world’s most capable, commercially‑scalable, and safe humanoid robot... Show more

 • Promoted

Neural Network Performance Engineer

HumanoidGreater London, England, GB
Full-time

Here at Humanoid, we believe in a future where robots amplify human potential.That’s why we’ve set out on a mission to build the world’s most capable, commercially‑scalable, and safe humanoid robot... Show more

 • Promoted

Staff ML Engineer - Lead Data & NLP for AI Search (London)

SR2 | Socially Responsible Recruitment | Certified B CorporationGreater London, England, GB
Full-time

An innovative AI technology platform based in London seeks a Staff ML Engineer to lead Data and NLP projects.This role involves owning search performance, optimizing algorithms, and implementing la... Show more