Session 5: Data in the Age of Generative AI
Massive datasets are the cornerstone for developing large language models (LLMs) and other generative AI. However, these datasets have also sparked debates regarding generative AI, highlighted by several copyright disputes involving OpenAI. This talk explores critical aspects of data creation and attribution for generative AI. Throughout the project, the researchers aim to ground their research with real-world legal and policy considerations and high-impact applications in law and medicine. Learn more about other research supported by the Hoffman-Yee Research Grant program here: https://hai.stanford.edu/research/grant-programs/hoffman-yee-research-grants?section=2024-grant-recipients 00:00:00 Lecture 00:29:06 Q&A