Post by Anthropic on X
Anthropic@AnthropicAI
XPrevious research suggested that attackers might need to poison a percentage of an AI model’s training data to produce a backdoor.
Our results challenge this—we find that even a small, fixed number of documents can poison an LLM of any size.
Read more: t.co
191 likes12 repliesPosted Oct 9, 2025