Tag
#prompt injection
3 stories taggedprompt injection.

AI Security
OpenAI Built an AI That Hacks Its Own Models to Make Them Safer
GPT-Red is an automated red-teaming system that attacks OpenAI's own chatbots to find weak spots before real attackers do. It already discovered a trick that human testers had missed.
4 min read

AI Security
One Poisoned Email Can Secretly Rewrite Your AI Assistant's Memory
A newly described attack called MemGhost shows how a single message to your inbox can plant a false 'fact' inside an AI agent's long-term memory, without you ever knowing.
3 min read

AI Security
Researchers Are Turning Hackers' Favourite AI Weapon Against Them
A cybersecurity firm found that hiding special instructions inside cloud credentials can cause AI hacking tools to shut themselves down.
3 min read