Skip to content

Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback.

Bai, Y., Jones, A., Ndousse, K., Askell, A., Chen, A., DasSarma, N., Drain, D., et al. (2022). Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback.

Type
Preprint
Intellectual base
Grey literature
Year
2022
DOI
10.48550/arxiv.2204.05862
Link
https://doi.org/10.48550/arXiv.2204.05862
Cited from
systems_cybernetics

Cited by

1 citation across 1 artifact.

Each citation links to the sentence it supports in the citing article.

Primes

Verification

This reference passed the adversarial substantiation pipeline: it was checked to exist and to support the claim it is attached to. See how references were verified.

Registry ID ref:a00779131b56 · see in the full table