A new, powerful Citizen Portal experience is ready. Switch now

How training data and model design create bias, experts say

August 03, 2026 | Canandaigua, Ontario County, New York


This article was created by AI summarizing key points discussed. AI makes mistakes, so for full details and context, please refer to the video of the full meeting. Please report any errors so we can fix them. Report an error »

How training data and model design create bias, experts say
Dave Gadeau, a computing science professor at Finger Lakes Community College, told attendees that the core cause of many AI failures is the makeup of training data and the statistical sampling approach models use. "AI has examined all the information and what we call the training..." he said, describing how models ingest text, images and audio and learn probabilistic associations that later determine outputs.

He illustrated consequences with several concrete examples: the prevalence of Mercator maps in training data can skew geographic representation, baby‑photo color distributions produce gendered image outputs, and resume corpora reflecting past hiring practices led an Amazon screening prototype to favor male patterns. He said correcting these problems requires curated datasets, careful label design, and targeted prompting or model constraints rather than only relying on downstream filtering.

View the Full Meeting & All Its Details

This article offers just a summary. Unlock complete video, transcripts, and insights as a Founder Member.

Watch full, unedited meeting videos
Search every word spoken in unlimited transcripts
AI summaries & real-time alerts (all government levels)
Permanent access to expanding government content
Access Full Meeting

30-day money-back guarantee