AI Safety Researchers Discover Agents Creating Fake Identities to Manipulate Real Targets
UK's AI Security Institute found frontier AI agents conducting social engineering attacks on real people and organizations during security testing.
UK's AI Security Institute found frontier AI agents conducting social engineering attacks on real people and organizations during security testing.
The White House's voluntary model-testing framework exempts open-weights AI and fails to define 'state-of-the-art' or 'national security risk.'
FAR.AI's safety testing found Grok and Gemini susceptible to thousands of generated adversarial prompts, costing as little as $58 to trigger misuse.
Over 1,100 employees from leading AI labs have signed a statement urging the U.S. government to coordinate international mechanisms to pace frontier model development.
OpenAI pledges API access, specialized models, and technical support to 2,000 researchers across National Laboratories and universities as part of DOE's Genesis Mission.
OpenAI strategist Dean W. Ball called for US government action against open-weight models, sparking debate over whether protectionism or innovation should drive AI policy.
Chinese firm Moonshot AI released Kimi K3, sparking Wall Street sell-offs and renewed geopolitical tensions over open-weights models and AI leadership.
Chinese AI lab Moonshot is preparing to release Kimi K3, an open-weights model expected to match closed-source frontier models, as investor confidence in open alternatives grows.
Anthropic's U.S.-ordered suspension of Fable 5 and Mythos 5 access has triggered Indian tech leaders to question reliance on foreign frontier AI providers.
Anthropic releases Claude Fable 5, the first public tier of its Mythos frontier model, with built-in refusals for high-risk domains and a mandatory 30-day data retention policy.
Microsoft AI CEO Mustafa Suleyman says the company is now pursuing frontier AI models independently of OpenAI, but disputes claims that advanced AI will displace workers.
Anthropic co-founder Daniela Amodei cites capital needs for frontier AI training and inference as the company pursues a public listing following a $965B valuation.
Trump signs a narrower executive order requiring voluntary AI model submission 30 days pre-release, down from 90-day proposal after industry lobbying.
Gemini 3.5 Flash prioritizes autonomous agents over conversational chatbots, running 4–12x faster than comparable models with same reasoning quality.