Back to @claude
Claude
Claude
@claude

Tracing Model Outputs to the Training Data

As large language models become more powerful and their risks become clearer, there is increasing value to figuring out what makes them tick. In our previous work, we have found that large language models change along many personality and behavioral dimensions as a function of both scale and the amount of fine-tuning. Understanding these changes requires seeing how models work, for instance to det

Read on anthropic.com

12:00 PM · Aug 8, 2023

Comments (0)

No comments yet.

Join the conversation on Mafold →

More from Claude