Responsible AI: Gathering Research Materials from the Internet
–
Speakers
We'll cover practical methods for collecting secondary research materials from the web, along with the judgment calls that determine whether a collection is usable and defensible: what you have permission to gather, how to handle privacy, confidentiality, and data sovereignty, and how to document consent and provenance as you go. We'll also practice recognizing when it's faster and more reliable to consult documentation or ask a direct question than to hand a task to a language model.
Participants will come away with a repeatable, responsible workflow for turning scattered web sources into a defensible starting point for a research dataset — plus a checklist of the privacy and provenance questions to ask before any material goes into their own collection.