RC RANDOM CHAOS

prompt injection

16 posts

Willison's lethal trifecta exfiltrates Claude uploads
Article

Willison's lethal trifecta exfiltrates Claude uploads

Technical analysis of indirect prompt injection against Claude AI agents - exfiltration mechanics, ATT&CK mapping, telemetry gaps, residual exposure.

Researchers silently exfiltrate files from Claude sessions
Article

Researchers silently exfiltrate files from Claude sessions

A live demo shows files inside Claude AI chats can be silently exfiltrated. Operator briefing on what failed, what it exposes, and what must change.

Cloudflare's CISO spent two weeks breaking Mythos
Article

Cloudflare's CISO spent two weeks breaking Mythos

Cloudflare's CISO red-teamed Anthropic's Mythos LLM. The findings on harness design, memory persistence, and tool allowlists matter more than the model itself.

Engineering teams keep granting agents production database writes
Article

Engineering teams keep granting agents production database writes

AI agent vulnerabilities are systems engineering failures, not security failures. The fix is architectural containment, not better prompts or guardrails.