OpenAI Built GPT-Red To Attack Its Own Models Before Prompt Hackers Do by Paul Balo July 16, 2026 0 OpenAI has introduced GPT-Red, an internal automated red-teaming system built to find prompt injection weaknesses and make AI agents safer ...
Context Bombing Turns Prompt Injection Into A Defence Against AI Hackers by Paul Balo July 14, 2026 0 Security researchers are testing context bombs, defensive prompt-injection traps designed to stop malicious AI agents during cyberattacks.