Skip to content
AI Lehel Briefing
← Back to latest
Models & research OpenAI

Aligning language models to follow instructions

We’ve trained language models that are much better at following user intentions than GPT-3 while also making them more truthful and less toxic, using techniques developed through our alignment research. These InstructGPT models, which are trained with humans in the loop, are now deployed as the default language models on our API.

Excerpt supplied by the publisher’s feed

Read the original article at OpenAI Opens in a new tab