Skip to main content

NLP Performance Engineer at gresearch

Location
London, England
Type
Full-time
Level
mid
Posted

Description

G-Research is hiring an NLP Performance Engineer to join its NLP Engineering team in London, focusing on large-scale LLM inference performance. The role is a specialist Quantitative Developer position that involves profiling, benchmarking, and optimizing LLM inference workloads across GPU architectures, developing reference implementations and tooling, and collaborating with researchers and infrastructure teams to improve the efficiency, reliability, and scalability of NLP systems.

For job seekers

Ready to find a role that actually fits?

Upload your résumé, start a Job Search Thread, and let Metaintro rank real openings against your experience — then guide you from search to offer.

Match

Compare live roles against your current evidence.

Position

Turn proof projects into role-specific applications.

Improve

Use market feedback to keep the skill plan current.

Return to navigation