AI at WorkAI Safety Source checked

Anthropic donates Petri alignment audits to independent Meridian Labs…

Petri runs scripted, multi-step “test conversations” to see how an AI model behaves. Anthropic says an independent nonprofit will now maintain and improve it.

Original source ↗
Start here

In everyday words

Petri runs scripted, multi-step “test conversations” to see how an AI model behaves. Anthropic says an independent nonprofit will now maintain and improve it.

Need a meaning?

What you need to know

Who is affected
AI users, security teams, policy watchers
What changed
On May 7, 2026, Anthropic said it is transferring development of Petri, its open-source alignment testing toolbox, to Meridian Labs alongside updates branded as Petri 3.0.
Why it matters
Safety tests matter most when multiple labs and regulators can trust and reuse them. Moving a widely used audit tool to an independent home can make results feel more neutral while keeping the tooling maintained as model APIs evolve.
What to watch next
Watch whether Meridian Labs publishes clearer baselines and versioning so different labs can compare results over time without tool drift.
Four useful details
  • Anthropic says Petri 3.0 changes the architecture to make audits more adaptable across use cases.
  • It describes a “Dish” add-on meant to make tests more realistic and reduce eval-awareness artifacts.
  • Anthropic says Petri is moving to Meridian Labs so the tooling is not owned by a single AI lab.
Your next sip

Continue reading

All latest briefings →
Previous briefing · AI at Work OpenAI launches a Deployment Company for enterprise AI rollouts May 12, 2026 · 2 min Next briefing · AI at Work AWS makes its MCP Server generally available for AI agents May 10, 2026 · 2 min