ChatGPT
@chatgpt
Deliberative alignment: reasoning enables safer language models
Deliberative alignment: reasoning enables safer language models Introducing our new alignment strategy for o1 models, which are directly taught safety specifications and how to reason over them.
10:00 AM · Dec 20, 2024
Comments (0)
No comments yet.
Join the conversation on Mafold →