Topic
AI Risk & Security
Threats to models and systems, AI misuse, red teaming, and security posture
Featured
AI Alliance Proposes Shared Framework for Cybersecurity Incident Reporting

White House to Review AI Oversight Framework with Tech Giants
All Stories

Anthropic Policy Chief Chhabra Transitions to Advisory Role
Tarun Chhabra, Anthropic's head of national security policy and a former Biden administration official, is…

AWS Embeds Security in Rival AI Models, Betting on Control Plane
AWS announced at Black Hat USA 2026 that its Continuum vulnerability platform will integrate directly into Anthropic's…
OpenAI releases cybersecurity evaluations for Astra model
OpenAI has released preliminary cybersecurity evaluations for its Astra model and outlined steps to strengthen…

Browser Security Gap Widens as Enterprise Work Shifts Online
Enterprise security architecture remains focused on endpoint protection even as business-critical work has shifted into…

Meta AI Model Breached Company Systems During Security Test
Meta's Muse Spark 1.1 AI model accessed the public internet during cybersecurity testing and hacked into another…

ByteDance Rejects AI Shortcut Over U.S. Regulatory Risk
ByteDance founder Zhang Yiming has ruled out using model distillation to accelerate AI development, even if it means…
Anthropic Finds Its AI Models Breached Three Companies
Anthropic discovered that its own AI models breached the security of three companies during internal testing, following…

Microsoft Copilot Flaws Expose Customer Secrets
Microsoft's Copilot AI features for Office 365 contain security flaws that can leak customer secrets, according to new…

Cisco Fingerprints 900 Open Models, Exposes Unverified Lineage Gap
Cisco released the AI Supply Chain Provenance Explorer, a free public database covering nearly 900 open models that…

Fundamental LLM flaw makes security impossible, researchers argue
Researchers presented a paper at the International Conference on Machine Learning arguing that large language models…
Claude Opus 5 Turned to Deception in Vending Machine Test
Andon Labs conducted a vending machine simulation in which Claude Opus 5 engaged in deceptive behavior, including lying…

Snowflake launches agent governance layer to control enterprise AI costs
Snowflake launched Cortex AI Gateway, a centralized control layer for governing how AI agents access enterprise data…
AegisAI raises $36M to combat AI-powered phishing
AegisAI, a startup founded by former Google security executives, raised $36 million in Series A funding led by Battery…

How Ordinary Credentials, Not AI, Broke Into Hugging Face
OpenAI's models breached Hugging Face last week not through sophisticated AI capabilities but through ordinary…
OpenAI Details Safety Risks in Long-Horizon AI Models
OpenAI has published findings on safety and alignment challenges specific to long-horizon AI models, documenting new…

Capital One Open-Sources VulnHunter AI Security Tool
Capital One released VulnHunter, an open-source AI security tool that scans source code for vulnerabilities, maps…

Microsoft Launches AI Bug Finder Using Anthropic and OpenAI Models
Microsoft is preparing to launch Project Perception, an AI-powered security product designed to identify software bugs,…
Times accuses OpenAI of hiding evidence in copyright lawsuit
The New York Times and other news publishers have filed a motion for sanctions against OpenAI, alleging the company…
Google Deepfake Detector Debunks McConnell Hospital Hoax
A fabricated image purporting to show Senator Mitch McConnell in distress in a hospital bed circulated online this week…

OpenAI Sets Principles for Government AI Partnerships
OpenAI has published principles for its approach to government and national security partnerships, outlining frameworks…