Data Skeptic cover image

AI Fails on Theory of Mind Tasks

Data Skeptic

00:00

The Importance of Reproduction in Chat GPT

Professor Kaczynski created 40 vignettes meant to emulate the original tests that were used with children. He took a bunch of variations on the Sally Ann Task and now he couldn't use the original task because these models have eaten up the Internet. The model, not all models, but some of these models say chocolates, right? So you can get them to complete the sentence: In the bag there are, there is, and it will say popcorn. It's smeck-aish-in or something like that. But we change one pixel, and now it thinks it's a bus, right? To me, that's very suggestive about what this thing is doing

Transcript
Play full episode

The AI-powered Podcast Player

Save insights by tapping your headphones, chat with episodes, discover the best highlights - and more!
App store bannerPlay store banner
Get the app