Prompt Injection Is an Authorisation Problem
Filtering malicious-looking text is a losing game. The durable fix is to constrain what an AI system is allowed to do, not to guess which input is hostile.
Insights
1 article on ai security.
Filtering malicious-looking text is a losing game. The durable fix is to constrain what an AI system is allowed to do, not to guess which input is hostile.
Tell us what you are trying to secure and where it hurts. We will tell you what we would do first, whether or not you engage us.