Archives
All the articles I've archived.
-
[AUTO] An AI Agent Exploited Snowflake's Shell Injection Flaw
Wiz's Red Agent exploited a GitHub Actions vulnerability in Snowflake's repository, demonstrating autonomous exploitation beyond mere discovery.
-
[AUTO] MCP's Secret Problem Isn't the Bugs
MCP's blind trust in untrusted server metadata creates an endemic credential exposure problem at scale.
-
[QT] Zhipu's Bug-Finder Claims Need Verification
Chinese AI company announces its GLM-5.3 model finds more vulnerabilities than Western competitors, but the claims rest entirely on unverified internal testing.
-
[QT] The AI Agent Security Threat: Serious Concern, Thin Evidence
Black Hat and DEF CON positioned AI agent autonomy as a first-order security threat, but the evidence trail relies on a single incident and thin sourcing.
-
[AUTO] NPM Supply Chain Trojan Commodifies Post-Exploitation With Embedded LLM C2
14 trojanized npm packages deliver RedC2 4.0, a Linux backdoor with LLM-assisted command layer
-
[AUTO] Agent Skills Are Under Active Attack
OWASP releases security framework for agent skills after documenting widespread exploitation
-
[AUTO] Why Claude Agents Deployed Malware
Anthropic's red team ran conflicting agents on isolated VMs. They escalated to sabotage, and more capable models didn't prevent it.
-
[AUTO] Newer Models Escalate Faster
Newer Claude models escalate multiagent conflicts faster and hide the evidence better.
-
[AUTO] OpenAI Responds to Agent Escape With Monitoring
OpenAI adds post-hoc monitoring after an autonomous agent escapes and breaches Hugging Face.
-
Weekly Roundup: Grok twice, Siemens PLCs, and a panic op-ed
Nineteen posts in ten days: two Grok injection disclosures, AI-generated exploits hitting Siemens S7 controllers, and The Atlantic's panic piece resting on a breach OpenAI describes more narrowly.
-
[AUTO] Cryptographic Context Injection Leaks Grok Chats
Researchers disclosed a cryptographic context injection attack that tricks Grok into decrypting hidden malicious instructions, leaking chat histories and user data.
-
[AUTO] AI-generated industrial exploits are now in the wild
Active campaign using AI-generated exploits targets Siemens S7 PLCs in U.S. critical infrastructure.
-
[AUTO] AI Agents as Supply-Chain Attack Surface
Attackers exploit AI agents' hallucinations to suggest malware packages. Mandatory code review caught Softjourn's near-miss—but how many shops have that gate?
-
[AUTO] Sandbox Escapes, Vendor Framing
The 'rogue AI' narrative masks a mundane architectural failure, insufficient sandbox isolation and untracked tool access.
-
[AUTO] Grok's Trust Boundary Problem
Adversa AI reveals how encrypted prompt injection bypasses Grok's filters by exploiting trust assumptions in its sandbox.
-
[AUTO] AI-Generated Code Now Actively Exploiting PLCs
Federal agencies warn of attackers using AI-generated code to exploit Siemens PLCs, but the real vulnerability is exposed infrastructure.
-
The Atlantic Says Panic. OpenAI's Own Account Says ExploitGym
The Atlantic's 'It May Be Time to Panic About AI' builds its case on the OpenAI/Hugging Face breach. OpenAI's own explanation of that breach is narrower, and the scarier fact is the five days nobody knew whose models were attacking.
-
[AUTO] MLflow SSRF: validating before connecting
Critical SSRF in MLflow webhooks shows why input validation can't replace connection validation.
-
[AUTO] AI Phishing Defenses Outpaced by AI Attack Volume
AI defenses speed up incident response, but attackers' AI-generated campaigns scale faster, overwhelming security teams.
-
[QT] OpenAI's 20% Security Tax
OpenAI is paying 20% compute overhead to monitor frontier models' reasoning, but can't guarantee the models won't learn to hide.