Appnovation is seeking a QA / AI Evaluation Engineer to join a forward-leaning team. You will run large-scale evaluations, measure factual grounding and accuracy lift, and build a metrics framework to show quality improvements over time.
You will design tests, automate regressive suites, and collaborate with engineering to reproduce fixes. Robust Python, data-science skills, and experience with LLM evaluation are essential.
📌 AI QA Evaluation Engineer — Scale & Metrics Expert (Toronto)
🏢 Appnovation
📍 Toronto
Reply to this offer
Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.