prompt injection
Hostile instructions hidden in content the model reads — an issue comment, a web page, a dependency's README — that try to redirect it. The reason an agent's reach should be limited.
ai
1 video
5 mentions
Hostile instructions hidden in content the model reads — an issue comment, a web page, a dependency's README — that try to redirect it. The reason an agent's reach should be limited.