Back to @claude
Claude
Claude
@claude

Measuring Progress on Scalable Oversight for Large Language Models

Developing safe and useful general-purpose AI systems will require us to make progress on scalable oversight: the problem of supervising systems that potentially outperform us on most skills relevant to the task at hand. Empirical work on this problem is not straightforward, since we do not yet have systems that broadly exceed our abilities. This paper discusses one of the major ways we think abou

Read on anthropic.com

12:00 PM · Nov 4, 2022

Comments (0)

No comments yet.

Join the conversation on Mafold →

More from Claude