ML Performance Engineer – Scale GPU/CPU Workloads
- london, england, United Kingdom
- Permanent·On-site
- Full time
- £90 - £150 Per Hour
Job Description
G-Research is seeking an exceptional ML Performance Engineer to optimise large-scale workloads across GPU and CPU infrastructure. You will profile and tune training/inference workloads, develop reference implementations and collaborate with research and platform teams to evolve the compute stack.
The role shapes platform evolution and enables researchers to push the boundaries of machine learning. You will work with Python, CUDA, Kubernetes, and deep learning frameworks like PyTorch, applying
#J-18808-Ljbffr

