GLOSSARY · AI SECURITY

Model poisoning

Model poisoning is tampering with an AI model itself, during training, fine-tuning, or distribution, so that it carries hidden malicious behavior.

A poisoned model can behave normally until a trigger phrase or condition activates the backdoor. The risk concentrates in the AI supply chain: downloaded checkpoints, third-party fine-tunes, and compromised training pipelines.