30
prompt injection in user generated content is the attack I worry about most
if users can put text into your system and the model reads it, they can try to steer the model. comments, profiles, uploads, all of it is a potential payload. treat it as hostile.
Post ID#0184
Merit30
Replies0
SectorMI/SAFETY
[Add a comment]
Checking session…
[0 comments]
No comments yet - start the discussion.