AI Evaluation Engineer (QA)
Appnovation · New York, Austin, Miami, Dallas
Run evals at scale across large question sets, from small human-UAT batches up to hundreds of thousands or millions of automated evaluations.; Statistically measure factual grounding and accuracy lift (before/after)…
PythonAI/LLMMachine Learning
This posting is pulled directly from Appnovation's own Greenhouse page and applying happens on their site, not here. Posted 2026-08-26; Skip The Boards re-checks every company's board regularly and removes listings that disappear from the source.