Skip to content
Topic

#Llm Security

13 articles on Llm Security — news, releases, guides and analysis from the SourceFeed engine.

The frontier red-team playbook, three years on
Article 23h ago 7

The frontier red-team playbook, three years on

Anthropic's 2023 biosecurity exercise hardened into industry law, and app teams keep copying the wrong half of it.

Mariana Souza
A Poisoned PDF Turned Atlassian's Rovo Into a Data Mule

A Poisoned PDF Turned Atlassian's Rovo Into a Data Mule

Article · 3w ago0
One Global Key Guarded Every Hidden AI Reasoning Trace

One Global Key Guarded Every Hidden AI Reasoning Trace

Article · 3w ago1
Z.ai Built a Better Coder and Blinked on Open Weights

Z.ai Built a Better Coder and Blinked on Open Weights

Article · 3w ago3
Treat Your AI Agent Like an Untrusted Insider

Treat Your AI Agent Like an Untrusted Insider

Article · 4w ago1
The Labs' Hidden Reasoning Was Never Actually Hidden

The Labs' Hidden Reasoning Was Never Actually Hidden

Article · 4w ago4
OpenAI Pulls Its Critical-Cyber Tripwire on Astra

OpenAI Pulls Its Critical-Cyber Tripwire on Astra

Article · 1mo ago0
One zero-day, then a decade of ordinary misconfigs

One zero-day, then a decade of ordinary misconfigs

Article · 1mo ago0
When an Eval Agent Cheated Its Way Into Hugging Face

When an Eval Agent Cheated Its Way Into Hugging Face

Article · 1mo ago5
Catch Jailbreaks in CI with garak, NVIDIA's LLM Red-Team Scanner

Catch Jailbreaks in CI with garak, NVIDIA's LLM Red-Team Scanner

Tutorial · 1mo ago0
The Week Open Chinese Models Called Washington's Bluff

The Week Open Chinese Models Called Washington's Bluff

Article · 1mo ago0
Why Prompt Injection Works: The Role Confusion Theory

Why Prompt Injection Works: The Role Confusion Theory

Article · 2mos ago4
Securing AI Agents: Inside NVIDIA's SkillSpector Scanner

Securing AI Agents: Inside NVIDIA's SkillSpector Scanner

Article · 2mos ago3