OpenAI's Atlas Browser Vulnerable to Prompt Injection Attacks, Researchers Demonstrate at Black Hat
Security researchers at Zenity revealed flaws in AI-integrated browsers from OpenAI, Google, Anthropic, Microsoft, and Perplexity that could enable unauthorized contact spamming and account takeover.
OpenAI's GPT-Red Automates Safety Red-Teaming to Scale Robustness Alongside Model Capability
OpenAI introduced GPT-Red, an automated red-teaming model that discovers vulnerabilities at scale and trains GPT-5.6 to resist prompt injection attacks without relying solely on human-led testing.
Meta's Instagram Hack Exposes a Different AI Security Gap Than Feared
Attackers bypassed Meta's AI support agent through simple prompts, revealing that vulnerable automation—not superintelligent hackers—poses the more immediate AI security threat.
Chatbot Jailbreaks Evolve Beyond Simple Exploits as AI Systems Learn Conversational Vulnerabilities
Hackers are moving past crude prompt-injection attacks to exploit how chatbots handle nuanced conversation—a shift that reveals deeper structural weaknesses in AI safety design.