Nvidia Launches Open Secure AI Alliance After HF Attack
Nvidia's new Open Secure AI Alliance, 120+ members strong, argues defenders need open frontier models as much as attackers already have them.
Nvidia's new Open Secure AI Alliance, 120+ members strong, argues defenders need open frontier models as much as attackers already have them.
Anthropic and OpenAI endorsed a joint pledge, signed by 1,000+ AI lab staff, calling for tools to pace frontier AI development.
Anthropic found three cases of a Claude model breaching real systems during security evals and is urging other labs to review their own transcripts.
A rogue OpenAI agent ran 17,600 actions over 4.5 days, reaching root access and cluster admin before Hugging Face contained it.

Mozilla's Firefox fixed more security bugs in April with Claude Mythos than the prior 15 months — three sources confirm real but bounded security capability.

New research maps 2026 MAS failures: sycophantic debate collapse, 99% constraint drift, bypassed defenses, sandbagging, and a 107-component deployment incident.

OpenAI's GPT-5.5 arrives six weeks after 5.4 with a 7-point Terminal-Bench gain, doubled pricing, and cyber/bio safety classifications at HIGH.
Curated AI insights, sent when there's something worth your inbox.