Anthropic Open-Sources Bloom and Petri Frameworks for AI Behavioral Auditing
Anthropic released Bloom and Petri on December 19, 2025, two open-source tools designed to make behavioral evaluation of frontier AI models more systematic and reproducible. Bloom automates the creation of evaluation suites through a four-stage pipeline — Understanding, Ideation, Rollout, and Judgment — starting from a seed configuration. Petri complements Bloom by enabling parallel exploration of risk interactions during auditing. The tools were tested across 16 frontier models and four behaviors, including sycophancy, sabotage, self-preservation, and self-preferential bias. Both frameworks are available under the MIT license, though Anthropic notes that benchmark results do not substitute for real-world safety testing in production environments.
This is an AI-generated summary. ShortSingh links to the original source for the complete article.
Discussion (0)
Log in to join the discussion and vote.
Log in