TrustML Young Scientist Seminar #83 20240902 Talk by Pang Wei Koh (University of Washington)

2024/9/25 10:46

説明

The 83rd Seminar

Date and Time: September 2, 2024: 10:30 am – 12:00 am (JST)
Venue: Online and Meeting Room 1 at the RIKEN AIP Nihonbashi office

Title:
Reliable data use: Synthesis, retrieval, and interaction

Speaker:
Pang Wei Koh (Assistant Professor, University of Washington)

Abstract:
How can we better use our data to build more reliable and responsible models? I will first discuss when it might be useful to train on synthetic image data derived, in turn, from a generative model trained on the available real data. Next, I will describe how scaling up the datastore for retrieval-based language models can significantly improve performance, indicating that the amount of data used at inference time—and not just at training time—should be considered as a new dimension of scaling language models. Finally, I will discuss how the static nature of most of our training data leads to language model failures in interactive settings.

Bio:
Pang Wei Koh is an assistant professor in the Allen School of Computer Science and Engineering at the University of Washington, a visiting research scientist at AI2, and a Singapore AI Visiting Professor. His research interests are in the theory and practice of building reliable machine learning systems. His research has been published in Nature and Cell, featured in media outlets such as The New York Times and The Washington Post, and recognized by the MIT Technology Review Innovators Under 35 Asia Pacific award and best paper awards at ICML and KDD. He received his PhD and BS in Computer Science from Stanford University. Prior to his PhD, he was the 3rd employee and Director of Partnerships at Coursera.

日曜日	月曜日	火曜日	水曜日	木曜日	金曜日	土曜日
		1日のイベントページへのリンク	2日のイベントページへのリンク	3日のイベントページへのリンク	4日のイベントページへのリンク	5日
6日	7日	8日	9日のイベントページへのリンク	10日のイベントページへのリンク	11日	12日
13日	14日	15日	16日のイベントページへのリンク	17日	18日	19日
20日	21日	22日	23日	24日	25日	26日
27日	28日	29日	30日	31日

革新知能統合研究センター

動画ライブラリ