
AI Fails on Theory of Mind Tasks
Data Skeptic
The Importance of Reproduction in Chat GPT
Professor Kaczynski created 40 vignettes meant to emulate the original tests that were used with children. He took a bunch of variations on the Sally Ann Task and now he couldn't use the original task because these models have eaten up the Internet. The model, not all models, but some of these models say chocolates, right? So you can get them to complete the sentence: In the bag there are, there is, and it will say popcorn. It's smeck-aish-in or something like that. But we change one pixel, and now it thinks it's a bus, right? To me, that's very suggestive about what this thing is doing
00:00
Transcript
Play full episode
Remember Everything You Learn from Podcasts
Save insights instantly, chat with episodes, and build lasting knowledge - all powered by AI.