We haven't solved how to prevent large language models from inadvertently memorizing and later leaking sensitive training data.
open
Global / Unspecified, Global
Even with privacy safeguards, AI models can sometimes reproduce specific pieces of sensitive information they encountered during training when prompted the right way. Fully preventing this kind of unintentional data leakage remains an unresolved technical challenge.
Citation ID: WS00774
Title: We haven't solved how to prevent large language models from inadvertently memorizing and later leaking sensitive training data.
URL: https://worldsolve.org/index.php?api=problem&id=774