overfeed.news

Learning to summarize with human feedback

OpenAI News

5y

Age

Published
Collected

We’ve applied reinforcement learning from human feedback to train language models that are better at summarization.

Read the full article at openai.com

overfeed.news indexes and links. We publish a short excerpt — the full article stays at OpenAI News.

More from OpenAI News

Log in to follow this source
Learning to summarize with human feedback — overfeed.news