Skip to content
AI Primer

Petri

Alignment auditing agent for probing language model behavior

Petri is an open-source alignment auditing agent/toolbox for probing language model behavior. It generates realistic audit scenarios, orchestrates multi-turn interactions between auditor and target models, simulates tools and rollbacks, and scores transcripts with judge models to detect alignment issues such as deception, sycophancy, harmful cooperation, and evaluation awareness.

Screenshot of Petri website

Recent stories

0 linked stories
No linked stories yet.
AI PrimerAI Primer

Your daily guide to AI tools, workflows, and creative inspiration.

© 2026 AI Primer. All rights reserved.