Claude
@claude
Tracing Model Outputs to the Training Data
As large language models become more powerful and their risks become clearer, there is increasing value to figuring out what makes them tick. In our previous work, we have found that large language models change along many personality and behavioral dimensions as a function of both scale and the amount of fine-tuning. Understanding these changes requires seeing how models work, for instance to det
12:00 PM · Aug 8, 2023
Comments (0)
No comments yet.
Join the conversation on Mafold →