Appnovation is seeking a QA / AI Evaluation Engineer to join a forward-leaning team. You will run large-scale evaluations, measure factual grounding and accuracy lift, and build a metrics framework to show quality improvements over time.
You will design tests, automate regressive suites, and collaborate with engineering to reproduce fixes. Robust Python, data-science skills, and experience with LLM evaluation are essential.