Evaluating Large Language Models in Theory of Mind Tasks.¶
Kosinski, M. (2024). Evaluating Large Language Models in Theory of Mind Tasks. Proceedings of the National Academy of Sciences, 121(45).
Cited by¶
1 citation across 1 artifact.
Each citation links to the sentence it supports in the citing article.
Primes¶
- Theory Of Mind
- False-belief tasks and Sally–Anne protocols have been ported directly into evaluations of large language models and embodied agents, revealing analogous failure modes.
This sourcePorts false-belief and Sally–Anne-style protocols into evaluations of large language models, revealing analogous performance and failure patterns.
- False-belief tasks and Sally–Anne protocols have been ported directly into evaluations of large language models and embodied agents, revealing analogous failure modes.
Verification¶
This reference passed the adversarial substantiation pipeline: it was checked to exist and to support the claim it is attached to. See how references were verified.
Registry ID ref:ab91a936ef7b · see in the full table