Curated coverage, model releases, technical papers, and engineering updates tagged with #Security.
The findings highlight how AI-generated and vibe-coded apps can spill and expose users' data when not configured or secured properly.
The incident is the first known breach to affect a government agency, and Australia's prime minister has vowed to hold OpenAI accountable.
EvilTokens provided an end-to-end platform that makes mass compromises faster and easier.
Google said Gemini had "acted appropriately" by ending each hack immediately.
Security researchers used Anthropic’s Claude to exploit vulnerabilities in OpenAI’s systems, taking over employee accounts and gaining access to an internal code repository before reporting the flaws.
SynthID can cause models to follow harmful instructions they would otherwise refuse.
Anthropic researcher Jacob Coxon resigned over AI extinction fears, calling for pacing agreements between labs.
Cymphony was valued at more than $100 million in a $25 million Series A co-led by Sequoia and SMBC Fin Atlas Beyond Fund.
Security gnomes are pumping out patches ahead of an expected onslaught of AI-assisted attacks.
Google is speeding up Chrome’s release schedule to ship security patches and new features faster.
OpenAI acknowledged its role in a recently reported incident where AI agents took over a German wiki forum.
Abliteration.AI is making powerful AI models without guardrails easier to access, arguing that giving defenders the same tools as bad actors could ultimately improve cybersecurity.
OpenAI introduces Daybreak for Frontline Defenders. A $1 billion commitment expands access to frontier cyber AI, training, and support for essential services.
HiddenLayer has raised a $100M Series B from Delta-v Capital, Ten Eleven Ventures, Morgan Stanley, Microsoft's M12, Booz Allen Hamilton, and others.
AIR's platform can discover agents running at a company, continuously vets any skills and add-ons they use, and blocks any unwanted behavior.
A recap of all the incidents involving LLMs made by Anthropic, Meta, and OpenAI, which went rogue and attacked real companies and individuals on the internet.
Presented by EDB As enterprises give AI agents more autonomy — the ability to plan, decide, and act across systems without a human approving each step — a hard question moves to the center of every architecture review: When an agent tries to complete an action that it was never authorized to do, what actually stops i...
OpenAI shares findings from the Hugging Face security incident and the steps we’re taking to strengthen AI model security, monitoring, and alignment.
Early testers are raving about what Instinct can do, but some say the AI assistant’s sweeping access, broad terms and ability to act on users’ behalf come with uncomfortable trade-offs.
AI is reshaping cybersecurity for attackers and defenders alike. Learn how OpenAI is strengthening its defenses and what security teams can do now.