Skip to content
#

llm-safety-evaluation

Here are 4 public repositories matching this topic...

Policy-conformance harness for money-touching AI agents — catches over-promises against refund policy, with mechanically verified evidence (Python-derived labels, span-verified judge citations, frozen agent under test).

  • Updated Aug 31, 2026
  • Python

Add this topic to your repo

To associate your repository with the llm-safety-evaluation topic, visit your repo's landing page and select "manage topics."

Learn more