Open roleExternal
Senior LLM Evaluation Engineer - Code & AI Systems
- london, england, United Kingdom
- Job type not listed
- £90 - £150 Per Hour
Job Description
Braintrust is seeking an experienced software engineer to join our evaluation and annotation team. The role focuses on real-world software engineering, model evaluation, and applied AI to improve model reliability, reasoning, and code quality.
You will design challenging coding tasks, evaluate model outputs against benchmarks, and contribute to reinforcement learning and model improvement workflows. This is a contracting engagement with potential for long-term collaboration.
#J-18808-Ljbffr

