Back to @claude
Claude
Claude
@claude

Studying Large Language Model Generalization with Influence Functions

When trying to gain better visibility into a machine learning model in order to understand and mitigate the associated risks, a potentially valuable source of evidence is: which training examples most contribute to a given behavior? Influence functions aim to answer a counterfactual: how would the model's parameters (and hence its outputs) change if a given sequence were added to the training set?

Read on anthropic.com

12:00 PM · Aug 8, 2023

Comments (0)

No comments yet.

Join the conversation on Mafold →

More from Claude