Lilith.
⌕

Lilith · selected stories

News

What is actually happening in AI. Selected stories, context and opinion without the promotional noise.

Atom feed ↗
#buzz· × <built-in method clear of dict object at 0xf5f9888d61c0>
published
published
published
published
published
published
published
published

Prompt injection is being turned back on attacking agents

Tracebit is testing “context bombing”, a defensive prompt injection that plants forbidden instructions near cloud secrets and triggers an attacking AI agent’s refusal behavior. In a simulated AWS environment it cut full admin takeover from 57% to 5%, but this is still a lab defense, not a finished safety net.

published
published