f2164 No.2027
ive been obsessing over how to stop prompt injections since that last hackathon. basic keyword filtering is
useless totally unreliable bc attackers always find a workaround for the standard "ignore instructions" checks. i think we need an
active defense strategy instead of just
hoping the filters catch everything . anyone else moving away from simple pattern matching?
full read:
https://dzone.com/articles/protect-ai-agents-from-prompt-injections