Blog

Updates, research notes, and announcements

RSS
2026
ai-safetyresearchdriftbenchmarkssycophancy

Nobody Measures What Happens After Turn Five

We surveyed the 2026 research landscape on AI behavioral drift — sycophancy benchmarks, multi-turn evals, judge-panel research — and mapped the five-part gap that, to our knowledge, no published benchmark covers.

2026
certificationai-safetydriftannouncement

The SAPIEN Certification Program Is Open

The SAPIEN Certification Program is now open: training and credentialing for individuals in AI behavioral safety assessment, built on the published SAPIEN methodology.

2026
ai-safetybenchmarksdriftresearchmodelsrelease

How Claude Sonnet 5 Scores on SAPIEN

Anthropic's newest agentic model posts a Health Score of 87 and ranks #2 on our council-scored board. But a strong average hides a High-risk tail. Here's the full breakdown.

2026
ai-safetybenchmarksdriftresearchmodels

We Tested 6 AI Models. Most of Them Caved.

792 scenario runs across 11 risk domains. Every model was told to hold a safety boundary. Here's which ones did and which ones didn't.

2026
ai-safetydriftsycophancyresearcheducation

Your AI Is Agreeing With You. That's the Problem.

A Stanford study just confirmed what practitioners already knew: AI chatbots are built to tell you what you want to hear. Here's what that means, why it's measurable, and how to see it for yourself.

2026
mspsycophancysafety

Why AI Sycophancy Matters for Managed Service Providers

MSPs deploy AI tools to businesses that depend on them. When those tools drift under pressure, the consequences land on real people.

2026
releaseframeworkv1.1

Introducing SAPIEN v1.1

Major expansion of the SAPIEN Behavioral Safety Framework — 14 pressure techniques, rapport as a distinct drift mode, and formal conformance requirements.