
AI Fails on Theory of Mind Tasks
Data Skeptic
 00:00 
The Importance of Reproduction in Chat GPT
Professor Kaczynski created 40 vignettes meant to emulate the original tests that were used with children. He took a bunch of variations on the Sally Ann Task and now he couldn't use the original task because these models have eaten up the Internet. The model, not all models, but some of these models say chocolates, right? So you can get them to complete the sentence: In the bag there are, there is, and it will say popcorn. It's smeck-aish-in or something like that. But we change one pixel, and now it thinks it's a bus, right? To me, that's very suggestive about what this thing is doing
 Play episode from 34:37 
 Transcript 


