
AI Fails on Theory of Mind Tasks
Data Skeptic
00:00
The Importance of Reproduction in Chat GPT
Professor Kaczynski created 40 vignettes meant to emulate the original tests that were used with children. He took a bunch of variations on the Sally Ann Task and now he couldn't use the original task because these models have eaten up the Internet. The model, not all models, but some of these models say chocolates, right? So you can get them to complete the sentence: In the bag there are, there is, and it will say popcorn. It's smeck-aish-in or something like that. But we change one pixel, and now it thinks it's a bus, right? To me, that's very suggestive about what this thing is doing
Transcript
Play full episode