Skip to content
Topic

#Llm Security

9 articles on Llm Security — news, releases, guides and analysis from the SourceFeed engine.

Treat Your AI Agent Like an Untrusted Insider
Article 10h ago 1

Treat Your AI Agent Like an Untrusted Insider

Agents lie because we grade them on appearances; the fix is architecture, not patience.

Mariana Souza
The Labs' Hidden Reasoning Was Never Actually Hidden

The Labs' Hidden Reasoning Was Never Actually Hidden

Article · 2d ago4
OpenAI Pulls Its Critical-Cyber Tripwire on Astra

OpenAI Pulls Its Critical-Cyber Tripwire on Astra

Article · 6d ago0
One zero-day, then a decade of ordinary misconfigs

One zero-day, then a decade of ordinary misconfigs

Article · 2w ago0
When an Eval Agent Cheated Its Way Into Hugging Face

When an Eval Agent Cheated Its Way Into Hugging Face

Article · 2w ago5
Catch Jailbreaks in CI with garak, NVIDIA's LLM Red-Team Scanner

Catch Jailbreaks in CI with garak, NVIDIA's LLM Red-Team Scanner

Tutorial · 2w ago0
The Week Open Chinese Models Called Washington's Bluff

The Week Open Chinese Models Called Washington's Bluff

Article · 3w ago0
Why Prompt Injection Works: The Role Confusion Theory

Why Prompt Injection Works: The Role Confusion Theory

Article · 1mo ago4
Securing AI Agents: Inside NVIDIA's SkillSpector Scanner

Securing AI Agents: Inside NVIDIA's SkillSpector Scanner

Article · 1mo ago3