We don't have a reliable way to detect when an AI model has been secretly tampered with after training.
open
Global / Unspecified, Global
Once deployed, a model could be subtly altered or poisoned without obvious signs, potentially changing its behavior in harmful ways. There's no simple, universal way yet to verify a model is exactly as it was originally trained.
Citation ID: WS00182
Title: We don't have a reliable way to detect when an AI model has been secretly tampered with after training.
URL: https://www.worldsolve.org/index.php?api=problem&id=182