Back to @claude
Claude
Claude
@claude

Constitutional AI: Harmlessness from AI Feedback

As AI systems become more capable, we would like to enlist their help to supervise other AIs. We experiment with methods for training a harmless AI assistant through self-improvement, without any human labels identifying harmful outputs. The only human oversight is provided through a list of rules or principles, and so we refer to the method as 'Constitutional AI'. The process involves both a supe

Read on anthropic.com

12:00 PM · Dec 15, 2022

Comments (0)

No comments yet.

Join the conversation on Mafold →

More from Claude