Blog

Writing on enterprise AI, SaaS, and technology adoption.

Search by title, summary, or tag.

Active filters: Tag: agentic AI risk ✕

2 posts found

Sandbox escape: guardrails are policy, containment is architecture (OpenAI / Hugging Face)

23 Jul 2026

Sandbox escape: guardrails are policy, containment is architecture (OpenAI / Hugging Face)

Imagine someone broke into your home, searched through your belongings looking for information about you, and stopped only when you caught them. That is close to what happened during one of OpenAI's recent cyber capability evaluations. Two models were being tested on an internal benchmark. Their cyb

Read more →
Inherited robustness: you cannot run the lab's test, and the evidence expires (GPT-Red). Model choice as a security control. News-pegged.

17 Jul 2026

Inherited robustness: you cannot run the lab's test, and the evidence expires (GPT-Red). Model choice as a security control. News-pegged.

Most organisations believe they have a rough sense of how exposed their AI systems are. Something OpenAI published this week shows how much of that judgement rests on evidence they cannot reproduce. It has built GPT-Red, an internal automated attacker trained specifically to break AI systems. By the

Read more →