paper
referenced-only
2022
paper:l-training-language-models-to-follow-instr-2022

Training language models to follow instructions with human feedback

ByL. Ouyang·J. Wu·X. Jiang·D. Almeida·C. Wainwright·P. Mishkin+4 more

Related work— refs + corpus + external arXiv

Cited / in-corpus / arXiv badges show which signals surfaced each row. Multi-source rows weighted higher.

Similar preprints — Semantic Scholar

Cited by (5)