Putting AI to the Test: How to Benchmark, Verify, and Validate AI for Food Safety
In this hands-on session, participants will apply a structured framework to benchmark AI performance, verify AI-generated outputs, and compare human and AI-assisted approaches to food safety tasks.
9:45 AM - 10:30 AMFri
Food Safety
Speakers
Chenhao (Luke) Qian, Ph.D.
post-doc in the Department of Food Science
Cornell University
David Hatch
Chief Research Officer; Co-Founder
Food Safety Tech Research (FSTR)
Abigail Snyder, Ph.D.
Associate Professor of Food Science
Cornell University
How can food safety professionals determine whether an AI tool is accurate, reliable, and fit for its intended use? In this hands-on session, participants will apply a structured framework to benchmark AI performance, verify AI-generated outputs, and compare human and AI-assisted approaches to food safety tasks. Working in groups, participants will compare general-purpose and purpose-built AI tools and assess the quality and reliability of their outputs. Facilitated discussions will also explore practical implementation needs and barriers. Participants will leave with a practical understanding of how benchmarking and verification support fit-for-purpose validation and defensible use.