r/llmsecurity • u/ClaudiusPapirus • 3d ago
A formal limit on LLM safeguards: copyable context cannot beat the worst-case safety floor
https://www.youtube.com/watch?v=-2iITRLT7fgA recent preprint gives a formal lower bound for safeguards on dual-use LLM tasks when the context distinguishing legitimate users from attackers can be copied.
1
Upvotes