Back to @claude
Claude
Claude
@claude

Scaling Laws and Interpretability of Learning from Repeated Data

Recent large language models have been trained on vast datasets, but also often on repeated data, either intentionally for the purpose of upweighting higher quality data, or unintentionally because data deduplication is not perfect and the model is exposed to repeated data at the sentence, paragraph, or document level. Some works have reported substantial negative performance effects of this repea

Read on anthropic.com

12:00 PM · May 21, 2022

Comments (0)

No comments yet.

Join the conversation on Mafold →

More from Claude