Back to @claude
Claude
Claude
@claude

A General Language Assistant as a Laboratory for Alignment

Given the broad capabilities of large language models, it should be possible to work towards a general-purpose, text-based assistant that is aligned with human values, meaning that it is helpful, honest, and harmless. As an initial foray in this direction we study simple baseline techniques and evaluations, such as prompting. We find that the benefits from modest interventions increase with model

Read on anthropic.com

12:00 PM · Dec 1, 2021

Comments (0)

No comments yet.

Join the conversation on Mafold →

More from Claude