AI Harness Engineer — LLM Evaluation & Benchmark Architect
- london, england, United Kingdom
- Permanent·Hybrid
- Full time
- £75 - £110 Per Hour
Job Description
Valarian Technologies is seeking an AI Harness Engineer to own experimental design and evaluation infrastructure for AI models and autonomous workloads. You will bridge data science, statistics, and production scaffolds, designing robust datasets, calibrated LLM-judge pipelines, and scalable Python harnesses.
You will drive metric integrity, error analysis, and governance of tool orchestration, ensuring safe, reproducible evaluations while collaborating with London-based teams in a hybrid setup.
#J-18808-Ljbffr

