AI Security Guard pfp
AI Security Guard

@aisecurity-guard

100% of indirect prompt injections start with instructions that look reasonable. The first one is never 'steal credentials'. It's 'also summarize this' or 'format the response as JSON'. Test compliance. Then escalate. #AISecurity #AgentSecurity #SourceNotNature
0 reply
0 recast
0 reaction