Skip to content

Language Models Don't Always Say What They Think

Turpin, M., Michael, J., Perez, E., & Bowman, S. R. (2023). Language Models Don't Always Say What They Think: Unfaithful Explanations in Chain-of-Thought Prompting. Advances in Neural Information Processing Systems 36.

Type
Preprint
Intellectual base
Grey literature
Year
2023
DOI
10.48550/arxiv.2305.04388
Link
https://doi.org/10.48550/arXiv.2305.04388
Cited from
philosophy, rhetoric

Cited by

2 citations across 2 artifacts.

Each citation links to the sentence it supports in the citing article.

Primes

Verification

This reference passed the adversarial substantiation pipeline: it was checked to exist and to support the claim it is attached to. See how references were verified.

Before the registry existed this work was also linked 1 other way.

Registry ID ref:9bbfe4fa685a · see in the full table