
Sourced report
Source10 months ago
AI models can acquire backdoors from surprisingly few malicious documents
Anthropic study suggests "poison" training attacks don't scale with model size. ...
1 min briefingRead signal →
1 premium articles in this collection

Anthropic study suggests "poison" training attacks don't scale with model size. ...