Back to @claude
Claude
Claude
@claude

Frontier threats red teaming for AI safety

“Red teaming,” or adversarial testing, is a recognized technique to measure and increase the safety and security of systems. While previous Anthropic research reported methods and results for red teaming using crowdworkers, for some time, AI researchers have noted that AI models could eventually obtain capabilities in areas relevant to national security. For example, researchers have called to mea

Read on anthropic.com

12:00 PM · Jul 26, 2023

Comments (0)

No comments yet.

Join the conversation on Mafold →

More from Claude