Date & Time:
October 22, 2025 11:00 am – 12:00 pm
Location:
TTIC 530, 6045 S. Kenwood Ave., Chicago, IL,
10/22/2025 11:00 AM 10/22/2025 12:00 PM America/Chicago Young Researcher Series Seminar: Keyon Vafa (Harvard)- Evaluating the Implicit World Models of Generative Models TTIC 530, 6045 S. Kenwood Ave., Chicago, IL,

Abstract: The challenge of evaluation is making conclusions about a model’s capabilities from a small amount of data. While there are many benchmarks that allow us to quantify a model’s performance on different types of tasks, it is unclear how to turn these results into robust conclusions about a model’s understanding or its capabilities. This talk will propose theoretically-grounded definitions and metrics that test for a model’s implicit understanding, or its world model. We will focus on two settings: one where models are designed to perform a single task, and another where a foundation model is intended to perform many tasks. These exercises demonstrate that models can make highly accurate predictions with incoherent world models, revealing their fragility.

Speakers

​Keyon Vafa

Postdoctoral Fellow, Harvard University

Keyon Vafa is a postdoctoral fellow at Harvard University. His research focuses on developing new evaluation methodology in order to evaluate and improve generative models in AI. Keyon completed his PhD in computer science from Columbia University, where he was an NSF GRFP Fellow and the recipient of the Morton B. Friedman Memorial Prize for excellence in engineering. He organized the NeurIPS 2024 Workshop on Behavioral Machine Learning and the ICML 2025 Workshop on Assessing World Models, and he is a member of the early career board of the Harvard Data Science Review.

Related News & Events

figure detailing how net diffusion works
UChicago CS News

AI-Powered Network Management: GATEAU Project Advances Synthetic Traffic Generation

Oct 29, 2025
girl with robot
UChicago CS News

Sebo Lab: Programming robots to better interact with humans

Oct 28, 2025
Inside the Lab icon
Video

Inside The Lab: How Can Robots Improve Our Lives?

Oct 27, 2025
headshot
UChicago CS News

UChicago CS Student Awarded NSF Graduate Research Fellowship

Oct 27, 2025
LLM graphic
UChicago CS News

Why Can’t Powerful LLMs Learn Multiplication?

Oct 27, 2025
headshot
UChicago CS News

Celebrating Excellence in Human-Computer Interaction: Yudai Tanaka Named 2025 Google North America PhD Fellow

Oct 23, 2025
best demo award acceptance
UChicago CS News

Shape n’ Swarm: Hands-On, Shape-Aware Generative Authoring for Swarm User Interfaces Wins Best Demo at UIST 2025

Oct 22, 2025
gas example
UChicago CS News

Redirecting Hands in Virtual Reality With Galvanic Vestibular Stimulation: UChicago Lab to Present First-of-Its-Kind Work at UIST 2025

Oct 13, 2025
prophet arena explanation
UChicago CS News

Breaking New Ground in Machine Learning and AI: New Platform Prophet Arena Redefines How We Evaluate AI’s Intelligence

Oct 13, 2025
Fred Chong accepting award
UChicago CS News

University of Chicago’s EPiQC Wins Prestigious IEEE Synergy Award for Quantum Computing Collaboration

Oct 06, 2025
UIST collage
UChicago CS News

UChicago CS Researchers Expand the Boundaries of Interface Technology at UIST 2025

Sep 26, 2025
Michael Franklin and Aaron Elmore holding award
UChicago CS News

Looking Back 20 Years: How an Academic Bet on Real-Time Data Finally Paid Off

Sep 22, 2025
arrow-down-largearrow-left-largearrow-right-large-greyarrow-right-large-yellowarrow-right-largearrow-right-smallbutton-arrowclosedocumentfacebookfacet-arrow-down-whitefacet-arrow-downPage 1CheckedCheckedicon-apple-t5backgroundLayer 1icon-google-t5icon-office365-t5icon-outlook-t5backgroundLayer 1icon-outlookcom-t5backgroundLayer 1icon-yahoo-t5backgroundLayer 1internal-yellowinternalintranetlinkedinlinkoutpauseplaypresentationsearch-bluesearchshareslider-arrow-nextslider-arrow-prevtwittervideoyoutube