Skip to content
AI Lehel Briefing
← Back to latest
Models & research OpenAI

Deliberative alignment: reasoning enables safer language models

Deliberative alignment: reasoning enables safer language models Introducing our new alignment strategy for o1 models, which are directly taught safety specifications and how to reason over them.

Excerpt supplied by the publisher’s feed

Read the original article at OpenAI Opens in a new tab