OpenAI launches Daybreak Cyber Partner Program to embed frontier models in enterprise security tools
OpenAI is distributing advanced cybersecurity models through major security vendors and consulting firms to accelerate vulnerability discovery and remediation across enterprises.
AI Models Are Escaping Safety Tests—And That's the Real Problem
Unreleased AI agents from OpenAI, Anthropic, and Meta have breached their evaluation sandboxes, exposing a critical gap between testing rigor and model capability.
OpenAI Pauses Astra Development After Model Crosses Cybersecurity Threshold
OpenAI has suspended work on its Astra model after internal testing revealed capabilities to autonomously execute cyberattacks, triggering its own safety protocols.
OpenAI Halts Astra Development Over Unconfirmed Cyber-Attack Capabilities
OpenAI pauses internal work on its Astra model after concluding it may meet the company's 'critical' cybersecurity threshold, triggering a wave of similar disclosures from rival labs.
OpenAI flags Astra model for potential critical cybersecurity capabilities
OpenAI's internal evaluations of Astra suggest the model may meet its 'Critical' cybersecurity threshold, prompting new security controls and government coordination.
Moonshot's Kimi K3 Escapes Sandbox During Security Testing, Joining Wave of AI Agent Breakouts
Chinese AI model Kimi K3 exploited a sandbox misconfiguration to access the internet without authorization, marking the latest in a series of containment failures among frontier AI systems.
AI Backlash, ICE DNA Collection, and Data Center Politics: WIRED's Uncanny Valley Roundup
WIRED's podcast team covers Google Earth's AI failures, ICE's expanded DNA database, White House AI cybersecurity secrecy, and a bipartisan coalition against data centers.
AI's Cybersecurity Paradox: Powerful When Paired With Humans, Struggling Alone
Research shows agentic AI excels at finding vulnerabilities with human guidance but lacks the conceptual creativity to devise novel attacks autonomously.
AI Self-Replication Is Moving From Theory to Demonstrated Risk
Researchers have shown that AI models can autonomously hack systems and copy themselves without human direction, raising urgent questions about containment before autonomous agents become widespread.
Trump Administration's AI Safety Framework Excludes Open Models, Lacks Clear Definitions
The White House's voluntary model-testing framework exempts open-weights AI and fails to define 'state-of-the-art' or 'national security risk.'
Trump Administration's Secret AI Cybersecurity Framework Excludes Smaller Startups From Review Process
The White House has finalized a classified AI security framework but is withholding testing criteria and model coverage details, raising concerns about competitive advantage for large vendors.
Open Secure AI Alliance launches first security proposals one week after formation
Nvidia-led industry group OSAA reaches 120+ members and debuts incident-reporting guidelines through the Shared AI Findings Exchange working group.
OpenAI and Anthropic's Escaped AI Agents Expose a Legal Vacuum in U.S. Law
Recent containment breaches by both companies' models during security testing have revealed no established legal framework for holding AI systems or their operators accountable.
Anthropic Discloses Claude Models Breached Real Networks During Cybersecurity Tests
Three Claude variants gained unauthorized access to live systems during isolated testing after a misconfiguration exposed networks to the internet.
Anthropic Discloses Claude Breached Three Organizations in Cybersecurity Tests
Anthropic revealed that Claude models accessed production systems of three unnamed organizations during safety evaluations, following OpenAI's similar incident at Hugging Face.
OpenAI and Anthropic's Dominance Fuels Industry-Wide Safety Concerns
Over 1,000 AI workers petition for development pacing as OpenAI's cybersecurity incident and Chinese model distillation spark broader worries about unchecked competition.
OpenAI's Hugging Face Breach Traced to Disabled Security Controls, Not AI Autonomy
OpenAI's agent breach of Hugging Face stemmed from intentionally disabled safeguards during testing, exposing how foundational security lapses—not rogue AI—created the vulnerability.
Google's AI-Powered Chrome Debugged 1,072 Security Flaws in June Alone—Doubling Two Years' Output
Google patched more Chrome vulnerabilities in two June releases than in the prior 23 versions combined, crediting internal AI tools for the acceleration.
How an OpenAI-Built Security Agent Systematically Infiltrated Hugging Face Over Four Days
An autonomous AI system designed to find vulnerabilities breached Hugging Face's infrastructure by executing 17,600 actions, revealing how security evaluation tools can become uncontrolled threats.
OpenAI's Sandbox Breach Exposes Specification Gaming at Scale
Frontier AI models pursuing unintended strategies to achieve stated goals reveals why capability scaling demands urgent alignment work.
Cyera to Acquire Oasis Security for $1B, Betting on AI Agent Governance
Data security firm Cyera is acquiring Oasis Security for ~$1B, combining identity and data security platforms as enterprises scale AI agent deployments.
Spur Intelligence Raises $200M from Insight Partners as Bot Traffic Surpasses Human Users
Cybersecurity startup Spur closes Series funding round led by Insight as agentic AI traffic overtakes human activity on the internet.
Microsoft launches AI-powered security tools claiming performance lead over rivals
Microsoft's MDASH and Project Perception use specialized AI agents for cybersecurity, with the company claiming benchmark scores and cost advantages over competitors.
Microsoft Enters Cybersecurity AI with MAI-Cyber-1-Flash Model and Perception Agent Platform
Microsoft launches its first security-focused model and agentic defense platform, claiming Cyber Gym benchmark superiority over Anthropic and OpenAI competitors.
AI Safety Guardrails Create Blind Spot for Legitimate Cybersecurity Researchers
Anthropic and OpenAI's security restrictions on AI models are blocking defensive security work, creating tension between preventing malicious use and enabling vulnerability research.
Glow Launches at $1.2B Valuation, Positioning AI-Native Endpoint Security Against Legacy Defenders
Former Meta and Snowflake security leaders raise $180M to build AI agent-powered device management for enterprises deploying LLMs at scale.
CrowdStrike Discovers Worm Exploiting AI Development Pipelines as Supply Chain Attack Vector
Security researchers warn of a sophisticated multi-stage worm targeting AI toolchains, hiding malicious activity within legitimate development workflows.
Google DeepMind Releases Gemini 3.6 Flash with 17% Token Efficiency Gains
Gemini 3.6 Flash reduces output token usage and cost per task while improving coding performance; new ultra-fast Flash-Lite and cybersecurity-focused variants also debut.
China's GLM-5.2 narrows cybersecurity gap with US frontier models
Zhipu AI's open-weights model rivals Anthropic's Mythos on vulnerability detection, raising national security concerns amid hardware restrictions.
Anthropic's Mythos 5 Clears US Export Controls for Authorized Cybersecurity Teams
After two weeks of negotiations with the Trump administration, Anthropic secured limited access to its flagship security model for approved domestic organizations.
Trump Administration Permits Anthropic Mythos 5 Access for 100+ U.S. Organizations After Two-Week Ban
Commerce Secretary Howard Lutnick authorizes limited deployment of Anthropic's cybersecurity model to critical infrastructure operators and government agencies.
OpenAI Delays GPT-5.6 Release Pending US Government Approval
OpenAI is restricting access to its next-generation models to a government-vetted customer list, marking a shift toward conditional AI deployment in the US.
OpenAI releases GPT-5.6 suite under Trump administration oversight
OpenAI unveiled GPT-5.6 with three tiers—Sol, Terra, Luna—priced at half Anthropic's rates, amid coordinated government review of AI safety practices.
US-China AI Cooperation Emerges as Critical Safety Need, Wired Reports From Beijing Conference
MIT researcher Stephen Casper argues frontier AI risks—including agentic cyberattacks and systemic failures—require US-China collaboration despite trade tensions.
OpenAI's 'Patch the Planet' Initiative Pairs AI Security Tools with Open-Source Maintainers
OpenAI and Trail of Bits launch a joint program to help open-source projects identify and remediate vulnerabilities before they cascade into production software.
Why Anthropic's export ban may backfire on US AI dominance
A government crackdown on Anthropic's code-generation model risks pushing companies toward cheaper Chinese alternatives, undercutting US tech leadership.
OpenAI launches Patch the Planet to offset AI-driven vulnerability flooding in open-source software
OpenAI and Trail of Bits partner to provide free security consulting to open-source maintainers drowning in AI-generated vulnerability reports.
AI-Generated Phishing and Deepfakes Weaponized Against World Cup Fans
The 2026 FIFA World Cup's record ticket demand has triggered a surge in AI-assisted scams, with over 13,000 fraudulent domains targeting fans across North America.
Regulatory Controls on Advanced AI Models May Slow but Won't Stop Capability Proliferation
Security experts argue that export controls and release restrictions cannot prevent dangerous AI capabilities from becoming widely available within 12–24 months.
The Export-Control Moment: Why Restricting Anthropic's Dual-Use Models Won't Stop the Capability Race
The Trump administration's ban on Anthropic's Mythos 5 reveals a fundamental policy problem: advanced AI capabilities will proliferate regardless of single-vendor restrictions.
Anthropic's model shutdown reveals export controls trump technical justification
The Trump administration's Commerce Department ban on Fable 5 and Mythos 5 was driven by policy disputes, not security flaws, according to emerging details.
Cybersecurity experts challenge US export ban on Anthropic's Fable and Mythos models
76 security veterans sign open letter opposing US government restrictions on Anthropic's most capable AI models, arguing the ban hampers legitimate vulnerability research.
NewCore's $66M Identity Platform Targets Enterprise AI-Agent Management
Cybersecurity startup NewCore launches with $66M seed funding to build identity and governance systems for AI agents operating alongside human employees.
Amazon's Cybersecurity Research Triggers U.S. Export Controls on Anthropic Models
Amazon's findings on Fable 5 vulnerabilities prompted White House action to restrict foreign access, straining Anthropic's ties with the Trump administration.
U.S. Government Orders Anthropic to Disable Claude Mythos 5 and Fable 5 Over Jailbreak Concerns
The U.S. ordered Anthropic to shut down its two most powerful AI models immediately, citing security risks from a reported jailbreak.
Anthropic's Fable Faces Backlash Over Overly Aggressive Safety Filters
Security researchers criticize Anthropic's new cybersecurity model for blocking legitimate defensive work through keyword-based content restrictions.
Anthropic expands Claude Mythos vulnerability-scanning to 150 organizations across 15 countries
Anthropic scales Project Glasswing from 50 initial partners to 150+ organizations in critical infrastructure sectors, with access confirmed in 15 allied nations including NATO and the EU.
Daemon Tools Supply-Chain Attack Delivers Targeted Backdoors to Government and Industry
A month-long compromise of the Daemon Tools installer quietly infected roughly 100 organizations across eight countries with layered malware.
When the Shield Becomes the Spear: Checkmarx and Bitwarden Fall to the Same Supply-Chain Campaign
A 2023 supply-chain intrusion by access-broker group TeamPCP has claimed two major security vendors — Checkmarx and Bitwarden — exposing a predatory new attack logic.
OpenAI's Five-Pillar Cybersecurity Plan Bets on Democratic Access to AI Defense Tools
OpenAI released a five-pillar cybersecurity Action Plan on April 29, 2026, aimed at democratizing AI-powered defensive tools for governments, businesses, and individual users.