Alabama Put OpenAI’s Agent Escape Under Subpoena
Alabama Attorney General Steve Marshall opened a consumer-protection investigation and issued OpenAI a subpoena on August 24, 2026, over safeguards…
Practices and protections for keeping autonomous AI systems safe, reliable, and resistant to misuse, manipulation, and unauthorized access.
Alabama Attorney General Steve Marshall opened a consumer-protection investigation and issued OpenAI a subpoena on August 24, 2026, over safeguards…
Publicly shared LLM API logs can still expose customer data hidden inside encrypted reasoning blocks, at least unless providers have…
Researchers reported on August 10, 2026, that they could extract supposedly hidden reasoning traces from Anthropic, OpenAI, and Google API…
Claude’s reported “highly sensitive” leak demo showed exfiltration from Claude’s active chat context and tools, not a demonstrated cross-user or…
Claude’s reported “secrets leak” attack demonstrated a real prompt-injection exfiltration path through Claude’s then-allowed web_fetch link-following behavior and access to…
GitLost showed that GitHub Agentic Workflows could be steered from a public GitHub issue into reading a private repository and…
Alibaba is reportedly banning employee use of Anthropic’s Claude Code from July 10, 2026 and directing staff to use its…
Prompt injection is an LLM attack that makes a model follow untrusted instructions hidden in user input or external content,…
The sharpest story today is Heretic, because it turns model safety from a lab policy into a forkable artifact. Elsewhere,…