More jobs:
AI Evaluation Engineer
Job in
Greater London, London, Greater London, W1B, England, UK
Listed on 2026-07-18
Listing for:
Gazelle Global
Full Time
position Listed on 2026-07-18
Job specializations:
-
Software Development
AI QA / Validation Engineer, AI Reliability/ Performance Engineer, AI Engineer (Applied/Software)
Job Description & How to Apply Below
Location:
London (Hybrid – 1 day onsite)
Contract:
6–12 Months
We're looking for an AI Evaluation Engineer to help deliver and optimise cutting‑edge Generative AI solutions within a large enterprise environment.
Key Responsibilities- Define and implement evaluation strategies for LLMs, RAG pipelines and AI agents
- Build automated evaluation pipelines and benchmarking frameworks
- Establish evaluation metrics and create test datasets
- Evaluate prompt quality, model performance and response accuracy
- Validate RAG knowledge grounding and conduct safety, risk and compliance testing
- Support human‑in‑the‑loop evaluation and continuous AI optimisation
- Strong experience with LLMs, RAG architectures and prompt engineering
- Python for data analysis and evaluation pipelines
- Experience with AI evaluation tools such as Deep Eval, Prompt Tools or similar
- Knowledge of NLP quality assessment, benchmarking and experimentation frameworks
- Understanding of Responsible AI principles
Note that applications are not being accepted from your jurisdiction for this job currently via this jobsite. Candidate preferences are the decision of the Employer or Recruiting Agent, and are controlled by them alone.
To Search, View & Apply for jobs on this site that accept applications from your location or country, tap here to make a Search:
To Search, View & Apply for jobs on this site that accept applications from your location or country, tap here to make a Search:
Search for further Jobs Here:
×